arXiv (Cornell University) · 2020 · 52 citations · 18 references
Convolutional Neural NetworkScene AnalysisEngineeringMachine LearningImage Sequence AnalysisImage AnalysisData SciencePattern RecognitionComputational GeometryConvolution LayerMachine VisionObject DetectionComputer ScienceDeep LearningMask PredictionMedical Image ComputingComputer VisionScene UnderstandingImage SegmentationInstance Segmentation
Instance segmentation is one of the fundamental vision tasks. Recently, fully convolutional instance segmentation methods have drawn much attention as they are often simpler and more efficient than two-stage approaches like Mask R-CNN. To date, almost all such approaches fall behind the two-stage Mask R-CNN method in mask precision when models have similar computation complexity, leaving great room for improvement. In this work, we achieve improved mask prediction by effectively combining instance-level information with semantic information with lower-level fine-granularity. Our main contribution is a blender module which draws inspiration from both top-down and bottom-up instance segmentation approaches. The proposed BlendMask can effectively predict dense per-pixel position-sensitive instance features with very few channels, and learn attention maps for each instance with merely one convolution layer, thus being fast in inference. BlendMask can be easily incorporated with the state-of-the-art one-stage detection frameworks and outperforms Mask R-CNN under the same training schedule while being 20% faster. A light-weight version of BlendMask achieves $ 34.2% $ mAP at 25 FPS evaluated on a single 1080Ti GPU card. Because of its simplicity and efficacy, we hope that our BlendMask could serve as a simple yet strong baseline for a wide range of instance-wise prediction tasks. Code is available at https://git.io/AdelaiDet
18
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren et al. · 2016 · 214.9K citations · Full text
Image Classification, Deep Neural Networks, Machine Vision +14
Kaiming He, Georgia Gkioxari, Piotr Dollár et al. · 2017 · 27.9K citations
Object Instance Segmentation, Scene Analysis, Machine Vision +13
Ross Girshick · 2015 · 27.2K citations
Image Classification, Convolutional Neural Network, Image Analysis +11
Focal Loss for Dense Object Detection
Tsung-Yi Lin, Priya Goyal, Ross Girshick et al. · 2017 · 24.4K citations
Image Classification, Convolutional Neural Network, Image Analysis +15
Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks
Shaoqing Ren, Kaiming He, Ross Girshick et al. · arXiv (Cornell University) · 2015 · 18.2K citations · Full text