Attention-guided Chained Context Aggregation for Semantic Segmentation

Tang, Quan; Liu, Fagui; Zhang, Tong; Jiang, Jun; Zhang, Yu

doi:10.1016/j.imavis.2021.104309

Full-text links:

Download:

Current browse context:

cs.CV

< prev | next >

new | recent | 2002

Change to browse by:

Computer Science > Computer Vision and Pattern Recognition

Title: Attention-guided Chained Context Aggregation for Semantic Segmentation

Authors: Quan Tang, Fagui Liu, Tong Zhang, Jun Jiang, Yu Zhang

(Submitted on 27 Feb 2020 (v1), last revised 21 May 2021 (this version, v4))

Abstract: The way features propagate in Fully Convolutional Networks is of momentous importance to capture multi-scale contexts for obtaining precise segmentation masks. This paper proposes a novel series-parallel hybrid paradigm called the Chained Context Aggregation Module (CAM) to diversify feature propagation. CAM gains features of various spatial scales through chain-connected ladder-style information flows and fuses them in a two-stage process, namely pre-fusion and re-fusion. The serial flow continuously increases receptive fields of output neurons and those in parallel encode different region-based contexts. Each information flow is a shallow encoder-decoder with appropriate down-sampling scales to sufficiently capture contextual information. We further adopt an attention model in CAM to guide feature re-fusion. Based on these developments, we construct the Chained Context Aggregation Network (CANet), which employs an asymmetric decoder to recover precise spatial details of prediction maps. We conduct extensive experiments on six challenging datasets, including Pascal VOC 2012, Pascal Context, Cityscapes, CamVid, SUN-RGBD and GATECH. Results evidence that CANet achieves state-of-the-art performance.

Comments:	Wrong numbers in Table 7 of version v3 have been corrected
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Journal reference:	Image and Vision Computing 2021
DOI:	10.1016/j.imavis.2021.104309
Cite as:	arXiv:2002.12041 [cs.CV]
	(or arXiv:2002.12041v4 [cs.CV] for this version)

Submission history

From: Quan Tang [view email]
[v1] Thu, 27 Feb 2020 11:26:56 GMT (519kb,D)
[v2] Sun, 17 Jan 2021 07:00:54 GMT (1064kb,D)
[v3] Tue, 19 Jan 2021 08:01:47 GMT (0kb,I)
[v4] Fri, 21 May 2021 03:25:20 GMT (1098kb,D)

Which authors of this paper are endorsers? | Disable MathJax (What is MathJax?)

Link back to: arXiv, form interface, contact.

> cs > arXiv:2002.12041

Download:

Current browse context:

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

Computer Science > Computer Vision and Pattern Recognition

Title: Attention-guided Chained Context Aggregation for Semantic Segmentation

Submission history