Timezone: »
We present SegNeXt, a simple convolutional network architecture for semantic segmentation. Recent transformer-based models have dominated the field of se- mantic segmentation due to the efficiency of self-attention in encoding spatial information. In this paper, we show that convolutional attention is a more efficient and effective way to encode contextual information than the self-attention mech- anism in transformers. By re-examining the characteristics owned by successful segmentation models, we discover several key components leading to the perfor- mance improvement of segmentation models. This motivates us to design a novel convolutional attention network that uses cheap convolutional operations. Without bells and whistles, our SegNeXt significantly improves the performance of previous state-of-the-art methods on popular benchmarks, including ADE20K, Cityscapes, COCO-Stuff, Pascal VOC, Pascal Context, and iSAID. Notably, SegNeXt out- performs EfficientNet-L2 w/ NAS-FPN and achieves 90.6% mIoU on the Pascal VOC 2012 test leaderboard using only 1/10 parameters of it. On average, SegNeXt achieves about 2.0% mIoU improvements compared to the state-of-the-art methods on the ADE20K datasets with the same or fewer computations.
Author Information
Meng-Hao Guo (Tsinghua University)

Third-year Ph.D candidate at the Department of Computer Science, Tsinghua University.
Cheng-Ze Lu
Qibin Hou (Nankai University)
Zhengning Liu (Tsinghua University, Tsinghua University)
Ming-Ming Cheng (Nankai University)
Shi-min Hu (Tsinghua University, Tsinghua University)
Related Events (a corresponding poster, oral, or spotlight)
-
2022 Poster: SegNeXt: Rethinking Convolutional Attention Design for Semantic Segmentation »
Dates n/a. Room
More from the Same Authors
-
2021 Poster: All Tokens Matter: Token Labeling for Training Better Vision Transformers »
Zi-Hang Jiang · Qibin Hou · Li Yuan · Daquan Zhou · Yujun Shi · Xiaojie Jin · Anran Wang · Jiashi Feng -
2020 Poster: ICNet: Intra-saliency Correlation Network for Co-Saliency Detection »
Wen-Da Jin · Jun Xu · Ming-Ming Cheng · Yi Zhang · Wei Guo -
2019 Poster: Unsupervised Scale-consistent Depth and Ego-motion Learning from Monocular Video »
Jiawang Bian · Zhichao Li · Naiyan Wang · Huangying Zhan · Chunhua Shen · Ming-Ming Cheng · Ian Reid -
2018 Poster: Self-Erasing Network for Integral Object Attention »
Qibin Hou · PengTao Jiang · Yunchao Wei · Ming-Ming Cheng