Semantic Segmentation With Context Encoding and Multi-Path Decoding

被引：128

作者：

Ding, Henghui ^{[1
]}

Jiang, Xudong ^{[1
]}

Shuai, Bing ^{[2
]}

Liu, Ai Qun ^{[1
]}

Wang, Gang ^{[3
]}

机构：

[1] Nanyang Technol Univ, Sch Elect & Elect Engn EEE, Singapore 639798, Singapore

[2] Amazon, Seattle, WA 98121 USA

[3] Alibaba AI Labs, Hangzhou 311121, Peoples R China

来源：

IEEE TRANSACTIONS ON IMAGE PROCESSING | 2020年 / 29卷

关键词：

Semantic segmentation; context encoding; gated sum; boundary delineation refinement; deep learning; CGBNet; convolutional neural networks; SCENE; FEATURES;

D O I：

10.1109/TIP.2019.2962685

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

Semantic image segmentation aims to classify every pixel of a scene image to one of many classes. It implicitly involves object recognition, localization, and boundary delineation. In this paper, we propose a segmentation network called CGBNet to enhance the segmentation performance by context encoding and multi-path decoding. We first propose a context encoding module that generates context-contrasted local feature to make use of the informative context and the discriminative local information. This context encoding module greatly improves the segmentation performance, especially for inconspicuous objects. Furthermore, we propose a scale-selection scheme to selectively fuse the segmentation results from different-scales of features at every spatial position. It adaptively selects appropriate score maps from rich scales of features. To improve the segmentation performance results at boundary, we further propose a boundary delineation module that encourages the location-specific very-low-level features near the boundaries to take part in the final prediction and suppresses them far from the boundaries. The proposed segmentation network achieves very competitive performance in terms of all three different evaluation metrics consistently on the six popular scene segmentation datasets, Pascal Context, SUN-RGBD, Sift Flow, COCO Stuff, ADE20K, and Cityscapes.

引用

页码：3520 / 3533

页数：14

共 83 条

[41] Learning Sparse High Dimensional Filters: Image Filtering, Dense CRFs and Bilateral Neural Networks [J].

Jampani, Varun ;

Kiefel, Martin ;

Gehler, Peter V. .

2016 IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2016, :4452-4461

[42]

Janoch A, 2011, 2011 IEEE INTERNATIONAL CONFERENCE ON COMPUTER VISION WORKSHOPS (ICCV WORKSHOPS)

[43]

Kendall Alex, WHAT UNCERTAINTIES W

[44] Convolutional Scale Invariance for Semantic Segmentation [J].

Kreso, Ivan ;

Causevic, Denis ;

Krapac, Josip ;

Segvic, Sinisa .

PATTERN RECOGNITION, GCPR 2016, 2016, 9796 :64-75

[45] Dynamic-structured Semantic Propagation Network [J].

Liang, Xiaodan ;

Zhou, Hongfei ;

Xing, Eric .

2018 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2018, :752-761

[46] Reversible Recursive Instance-level Object Segmentation [J].

Liang, Xiaodan ;

Wei, Yunchao ;

Shen, Xiaohui ;

Jie, Zequn ;

Feng, Jiashi ;

Lin, Liang ;

Yan, Shuicheng .

2016 IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2016, :633-641

[47] Graininess-Aware Deep Feature Learning for Pedestrian Detection [J].

Lin, Chunze ;

Lu, Jiwen ;

Wang, Gang ;

Zhou, Jie .

COMPUTER VISION - ECCV 2018, PT IX, 2018, 11213 :745-761

[48] RefineNet: Multi-Path Refinement Networks for High-Resolution Semantic Segmentation [J].

Lin, Guosheng ;

Milan, Anton ;

Shen, Chunhua ;

Reid, Ian .

30TH IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2017), 2017, :5168-5177

[49] Microsoft COCO: Common Objects in Context [J].

Lin, Tsung-Yi ;

Maire, Michael ;

Belongie, Serge ;

Hays, James ;

Perona, Pietro ;

Ramanan, Deva ;

Dollar, Piotr ;

Zitnick, C. Lawrence .

COMPUTER VISION - ECCV 2014, PT V, 2014, 8693 :740-755

[50] SIFT Flow: Dense Correspondence across Scenes and Its Applications [J].

Liu, Ce ;

Yuen, Jenny ;

Torralba, Antonio .

IEEE TRANSACTIONS ON PATTERN ANALYSIS AND MACHINE INTELLIGENCE, 2011, 33 (05) :978-994

← 1 2 3 4 5 6 7 8 9 →