IEEE Transactions on Artificial Intelligence · 2022 · 32 citations · 36 references
Convolutional Neural NetworkEngineeringMachine LearningReal-time Semantic SegmentationImage AnalysisData ScienceSemantic SegmentationRobot LearningVideo TransformerMachine VisionObject DetectionComputer EngineeringVision Language ModelComputer ScienceDeep LearningMedical Image ComputingNeural Architecture SearchComputer VisionSegmentation AccuracyObject RecognitionImage Segmentation
The architectural advancements in deep neural networks have led to remarkable leap-forwards across a broad array of computer vision tasks. Instead of relying on human expertise, neural architecture search (NAS) has emerged as a promising avenue toward automating the design of architectures. While recent achievements on image classification have suggested opportunities, the promises of NAS have yet to be thoroughly assessed on more challenging tasks of semantic segmentation. The main challenges of applying NAS to semantic segmentation arise from two aspects: 1) high-resolution images to be processed; 2) additional requirement of real-time inference speed (i.e., real-time semantic segmentation) for applications such as autonomous driving. To meet such challenges, we propose a surrogate-assisted multiobjective method in this article. Through a series of customized prediction models, our method effectively transforms the original NAS task to an ordinary multiobjective optimization problem. Followed by a hierarchical prescreening criterion for in-fill selection, our method progressively achieves a set of efficient architectures trading-off between segmentation accuracy and inference speed. Empirical evaluations on three benchmark datasets together with an application using Huawei Atlas 200 DK suggest that our method can identify architectures significantly outperforming existing state-of-the-art architectures designed both manually by human experts and automatically by other NAS methods. Code is available from here.
36
Deep Residual Learning for Image Recognition
Kaiming He, Xiangyu Zhang, Shaoqing Ren et al. · 2016 · 214.9K citations · Full text
Image Classification, Deep Neural Networks, Machine Vision +14
ImageNet: A large-scale hierarchical image database
Jia Deng, Wei Dong, Richard Socher et al. · 2009 IEEE Conference on Computer Vision and Pattern Recognition · 2009 · 60.2K citations
Densely Connected Convolutional Networks
Gao Huang, Zhuang Liu, Laurens van der Maaten et al. · 2017 · 43.3K citations
Geometric Learning, Convolutional Neural Network, Engineering +16