IEEE Transactions on Circuits and Systems for Video Technology · 2018 · 94 citations · 55 references
Convolutional Neural NetworkScene AnalysisEngineeringMachine LearningAttentionVideo InterpretationImage AnalysisPattern RecognitionVideo TransformerVision RecognitionMachine VisionVisual SaliencyObject DetectionDeep Learning TechniquesVideo UnderstandingDeep LearningComputer VisionScene InterpretationObject RecognitionSalient Object Detection
Recently, deep learning techniques have substantially boosted the performance of salient object detection in still images. However, the salient object detection in videos by using traditional handcrafted features or deep learning features is not fully investigated, probably due to the lack of sufficient manually labeled video data for saliency modeling, especially for the data-driven deep learning. This paper proposes a novel weakly supervised approach to the salient object detection in a video, which can learn a robust saliency prediction model by using very limited manually labeled data and a large amount of weakly labeled data that could be easily generated in a supervised approach. Furthermore, we propose a spatiotemporal cascade neural network architecture for saliency modeling, in which two fully convolutional networks are cascaded to evaluate the visual saliency from both spatial and temporal cues to lead the optimal video saliency prediction. The proposed approach is extensively evaluated on the widely used challenging data sets, and the experiments demonstrate that our proposed approach substantially outperforms the state-of-the-art salient object detection models.
55
Going deeper with convolutions
Christian Szegedy, Wei Liu, Yangqing Jia et al. · 2015 · 46.2K citations
Image Classification, Deep Neural Networks, Image Analysis +15
ImageNet Large Scale Visual Recognition Challenge
Olga Russakovsky, Jia Deng, Hao Su et al. · International Journal of Computer Vision · 2015 · 39.5K citations
Image Classification, Convolutional Neural Network, Machine Vision +7
Fully convolutional networks for semantic segmentation
Jonathan Long, Evan Shelhamer, Trevor Darrell · 2015 · 36.2K citations
Yangqing Jia, Evan Shelhamer, Jeff Donahue et al. · 2014 · 11.1K citations
Convolutional Neural Network, Machine Vision, Machine Learning +14