Journal Of Big Data · 2020 · 73 citations · 44 references
Attention RegionAnomaly DetectionMachine LearningEngineeringDeep Anomaly DetectionVideo SurveillanceVideo InterpretationVisual SurveillanceImage AnalysisData SciencePattern RecognitionVideo TransformerAnomaly EventsAnomaly BehaviorMachine VisionVideo UnderstandingDeep LearningComputer VisionNovelty Detection
Abstract This paper describes a method for learning anomaly behavior in the video by finding an attention region from spatiotemporal information, in contrast to the full-frame learning. In our proposed method, a robust background subtraction (BG) for extracting motion, indicating the location of attention regions is employed. The resulting regions are finally fed into a three-dimensional Convolutional Neural Network (3D CNN). Specifically, by taking advantage of C3D (Convolution 3-dimensional), to completely exploit spatiotemporal relation, a deep convolution network is developed to distinguish normal and anomalous events. Our system is trained and tested against a large-scale UCF-Crime anomaly dataset for validating its effectiveness. This dataset contains 1900 long and untrimmed real-world surveillance videos and splits into 950 anomaly events and 950 normal events, respectively. In total, there are approximately ~ 13 million frames are learned during the training and testing phase. As shown in the experiments section, in terms of accuracy, the proposed visual attention model can obtain 99.25 accuracies. From the industrial application point of view, the extraction of this attention region can assist the security officer on focusing on the corresponding anomaly region, instead of a wider, full-framed inspection.
44
Dropout: a simple way to prevent neural networks from overfitting
Nitish Srivastava, Geoffrey E. Hinton, Alex Krizhevsky et al. · 2014 · 34.2K citations
Learning Spatiotemporal Features with 3D Convolutional Networks
Du Tran, Lubomir Bourdev, Rob Fergus et al. · 2015 · 9.5K citations
Convolutional Neural Network, Image Analysis, Machine Vision +14
Bilateral filtering for gray and color images
Carlo Tomasi, Roberto Manduchi · 2002 · 8K citations
Large-Scale Video Classification with Convolutional Neural Networks
Andrej Karpathy, George Toderici, Sanketh Shetty et al. · 2014 · 6.3K citations
Convolutional Neural Network, Machine Vision, Machine Learning +14