Spatial Information Guided Convolution for Real-Time RGBD Semantic Segmentation

Chen, Lin-Zhuo; Lin, Zheng; Wang, Ziqin; Yang, Yong-Liang; Cheng, Ming-Ming

doi:10.1109/TIP.2021.3049332

Computer Science > Computer Vision and Pattern Recognition

arXiv:2004.04534 (cs)

[Submitted on 9 Apr 2020 (v1), last revised 8 Jan 2021 (this version, v2)]

Title:Spatial Information Guided Convolution for Real-Time RGBD Semantic Segmentation

Authors:Lin-Zhuo Chen, Zheng Lin, Ziqin Wang, Yong-Liang Yang, Ming-Ming Cheng

View PDF

Abstract:3D spatial information is known to be beneficial to the semantic segmentation task. Most existing methods take 3D spatial data as an additional input, leading to a two-stream segmentation network that processes RGB and 3D spatial information separately. This solution greatly increases the inference time and severely limits its scope for real-time applications. To solve this problem, we propose Spatial information guided Convolution (S-Conv), which allows efficient RGB feature and 3D spatial information integration. S-Conv is competent to infer the sampling offset of the convolution kernel guided by the 3D spatial information, helping the convolutional layer adjust the receptive field and adapt to geometric transformations. S-Conv also incorporates geometric information into the feature learning process by generating spatially adaptive convolutional weights. The capability of perceiving geometry is largely enhanced without much affecting the amount of parameters and computational cost. We further embed S-Conv into a semantic segmentation network, called Spatial information Guided convolutional Network (SGNet), resulting in real-time inference and state-of-the-art performance on NYUDv2 and SUNRGBD datasets.

Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2004.04534 [cs.CV]
	(or arXiv:2004.04534v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2004.04534
Related DOI:	https://doi.org/10.1109/TIP.2021.3049332

Submission history

From: Lin-Zhuo Chen [view email]
[v1] Thu, 9 Apr 2020 13:38:05 UTC (3,956 KB)
[v2] Fri, 8 Jan 2021 04:24:35 UTC (5,934 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Spatial Information Guided Convolution for Real-Time RGBD Semantic Segmentation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Spatial Information Guided Convolution for Real-Time RGBD Semantic Segmentation

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators