Semantic Segmentation With Context Encoding and Multi-Path Decoding

被引:120
|
作者
Ding, Henghui [1 ]
Jiang, Xudong [1 ]
Shuai, Bing [2 ]
Liu, Ai Qun [1 ]
Wang, Gang [3 ]
机构
[1] Nanyang Technol Univ, Sch Elect & Elect Engn EEE, Singapore 639798, Singapore
[2] Amazon, Seattle, WA 98121 USA
[3] Alibaba AI Labs, Hangzhou 311121, Peoples R China
关键词
Semantic segmentation; context encoding; gated sum; boundary delineation refinement; deep learning; CGBNet; convolutional neural networks; SCENE; FEATURES;
D O I
10.1109/TIP.2019.2962685
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Semantic image segmentation aims to classify every pixel of a scene image to one of many classes. It implicitly involves object recognition, localization, and boundary delineation. In this paper, we propose a segmentation network called CGBNet to enhance the segmentation performance by context encoding and multi-path decoding. We first propose a context encoding module that generates context-contrasted local feature to make use of the informative context and the discriminative local information. This context encoding module greatly improves the segmentation performance, especially for inconspicuous objects. Furthermore, we propose a scale-selection scheme to selectively fuse the segmentation results from different-scales of features at every spatial position. It adaptively selects appropriate score maps from rich scales of features. To improve the segmentation performance results at boundary, we further propose a boundary delineation module that encourages the location-specific very-low-level features near the boundaries to take part in the final prediction and suppresses them far from the boundaries. The proposed segmentation network achieves very competitive performance in terms of all three different evaluation metrics consistently on the six popular scene segmentation datasets, Pascal Context, SUN-RGBD, Sift Flow, COCO Stuff, ADE20K, and Cityscapes.
引用
收藏
页码:3520 / 3533
页数:14
相关论文
共 50 条
  • [1] Semantic segmentation network with multi-path structure, attention reweighting and multi-scale encoding
    Lin, Zhongkang
    Sun, Wei
    Tang, Bo
    Li, Jinda
    Yao, Xinyuan
    Li, Yu
    VISUAL COMPUTER, 2023, 39 (02): : 597 - 608
  • [2] Semantic segmentation network with multi-path structure, attention reweighting and multi-scale encoding
    Zhongkang Lin
    Wei Sun
    Bo Tang
    Jinda Li
    Xinyuan Yao
    Yu Li
    The Visual Computer, 2023, 39 : 597 - 608
  • [3] Encoding context and decoding aggregated information for semantic segmentation
    Zhang, Guodong
    Yang, Wenzhu
    Zhou, Guoyu
    COMPUTERS & GRAPHICS-UK, 2025, 126
  • [4] Multi-path Fusion Network For Semantic Image Segmentation
    Song, Hui
    Zhou, Yun
    Jiang, Zhuqing
    Guo, Xiaoqiang
    Yang, Zixuan
    2018 IEEE/CIC INTERNATIONAL CONFERENCE ON COMMUNICATIONS IN CHINA (ICCC), 2018, : 90 - 94
  • [5] Efficient Semantic Segmentation Using Multi-Path Decoder
    Bai, Xing
    Zhou, Jun
    APPLIED SCIENCES-BASEL, 2020, 10 (18):
  • [6] Multi-Path Spatial Detail-Aware Network for Semantic Segmentation
    School of Electronic Information Engineering, Shanxi Key Laboratory of Advanced Control and Equipment Intelligence, Taiyuan University of Science and Technology, Taiyuan, China
    Proc. - China Autom. Congr., CAC, (6330-6335):
  • [7] A Multi-Path Semantic Segmentation Network Based on Convolutional Attention Guidance
    Feng, Chenyang
    Hu, Shu
    Zhang, Yi
    APPLIED SCIENCES-BASEL, 2024, 14 (05):
  • [8] Context Encoding for Semantic Segmentation
    Zhang, Hang
    Dana, Kristin
    Shi, Jianping
    Zhang, Zhongyue
    Wang, Xiaogang
    Tyagi, Ambrish
    Agrawal, Amit
    2018 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR), 2018, : 7151 - 7160
  • [9] RefineNet: Multi-Path Refinement Networks for High-Resolution Semantic Segmentation
    Lin, Guosheng
    Milan, Anton
    Shen, Chunhua
    Reid, Ian
    30TH IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2017), 2017, : 5168 - 5177
  • [10] SEMANTIC SEGMENTATION WITH MULTI-PATH REFINEMENT AND PYRAMID POOLING DILATED-RESNET
    Cui, Zhipeng
    Zhang, Qiao
    Geng, Shijie
    Niu, Xiaoguang
    Yang, Jie
    Qiao, Yu
    2017 24TH IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING (ICIP), 2017, : 3100 - 3104