2018 · 12 citations · 15 references
Convolutional Neural NetworkEngineeringMachine LearningHuman Pose EstimationVideo ProcessingVideo SurveillanceVisual SurveillanceImage AnalysisData SciencePattern RecognitionCamera NetworkMachine VisionDepth-based NetworkObject DetectionHuman Body LandmarksComputer ScienceVideo UnderstandingDeep LearningComputer VisionPeople DetectionVideo Surveillance SystemsDeep-learning Approach
We propose a deep-learning approach for people detection on depth imagery. The approach is designed to be deployed as an autonomous appliance for identifying people attacks and intrusion in video surveillance scenarios. To this end, we propose a fully-convolutional and sequential network, named WatchNet, that localizes people in depth images by predicting human body landmarks such as head and shoulders. We use a large synthetic dataset to train the network with abundant data and generate automatic annotations. Adaptation to real data is performed via fine tuning with real depth images.The proposed method is validated in a novel and challenging database with about 29k top view images collected from several sequences including different people assaults. A comparative evaluation is given between our approach and other standard methods, showing remarkable detection results and efficiency. The network runs in 10 and 28 FPS using CPU and GPU, respectively.
15
Fully convolutional networks for semantic segmentation
Jonathan Long, Evan Shelhamer, Trevor Darrell · 2015 · 36.2K citations
Realtime Multi-person 2D Pose Estimation Using Part Affinity Fields
Zhe Cao, Tomas Simon, Shih-En Wei et al. · 2017 · 7.2K citations
Lokesh Boominathan, Srinivas S S Kruthiventi, R. Venkatesh Babu · 2016 · 517 citations
Water Filling: Unsupervised People Counting via Vertical Kinect Sensor
Xucong Zhang, Junjie Yan, Shikun Feng et al. · 2012 · 103 citations