A Survey on Deep Learning-based Architectures for Semantic Segmentation on 2D images

Ulku, Irem; Akagunduz, Erdem

doi:10.1080/08839514.2022.2032924

Computer Science > Computer Vision and Pattern Recognition

arXiv:1912.10230 (cs)

[Submitted on 21 Dec 2019 (v1), last revised 16 Mar 2022 (this version, v5)]

Title:A Survey on Deep Learning-based Architectures for Semantic Segmentation on 2D images

Authors:Irem Ulku, Erdem Akagunduz

View PDF

Abstract:Semantic segmentation is the pixel-wise labelling of an image. Since the problem is defined at the pixel level, determining image class labels only is not acceptable, but localising them at the original image pixel resolution is necessary. Boosted by the extraordinary ability of convolutional neural networks (CNN) in creating semantic, high level and hierarchical image features; several deep learning-based 2D semantic segmentation approaches have been proposed within the last decade. In this survey, we mainly focus on the recent scientific developments in semantic segmentation, specifically on deep learning-based methods using 2D images. We started with an analysis of the public image sets and leaderboards for 2D semantic segmentation, with an overview of the techniques employed in performance evaluation. In examining the evolution of the field, we chronologically categorised the approaches into three main periods, namely pre-and early deep learning era, the fully convolutional era, and the post-FCN era. We technically analysed the solutions put forward in terms of solving the fundamental problems of the field, such as fine-grained localisation and scale invariance. Before drawing our conclusions, we present a table of methods from all mentioned eras, with a summary of each approach that explains their contribution to the field. We conclude the survey by discussing the current challenges of the field and to what extent they have been solved.

Comments:	published in the J. of Applied Artificial Intelligence (09 Feb 2022)
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1912.10230 [cs.CV]
	(or arXiv:1912.10230v5 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1912.10230
Related DOI:	https://doi.org/10.1080/08839514.2022.2032924

Submission history

From: Erdem Akagündüz [view email]
[v1] Sat, 21 Dec 2019 09:31:09 UTC (4,097 KB)
[v2] Thu, 14 May 2020 15:05:12 UTC (4,098 KB)
[v3] Sat, 1 May 2021 19:40:57 UTC (4,525 KB)
[v4] Sat, 8 Jan 2022 15:18:51 UTC (4,477 KB)
[v5] Wed, 16 Mar 2022 10:28:03 UTC (4,478 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:A Survey on Deep Learning-based Architectures for Semantic Segmentation on 2D images

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:A Survey on Deep Learning-based Architectures for Semantic Segmentation on 2D images

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators