Bridging the Reality Gap for Pose Estimation Networks using Sensor-Based Domain Randomization

Hagelskjaer, Frederik; Buch, Anders Glent

Computer Science > Computer Vision and Pattern Recognition

arXiv:2011.08517 (cs)

[Submitted on 17 Nov 2020 (v1), last revised 17 Aug 2021 (this version, v3)]

Title:Bridging the Reality Gap for Pose Estimation Networks using Sensor-Based Domain Randomization

Authors:Frederik Hagelskjaer, Anders Glent Buch

View PDF

Abstract:Since the introduction of modern deep learning methods for object pose estimation, test accuracy and efficiency has increased significantly. For training, however, large amounts of annotated training data are required for good performance. While the use of synthetic training data prevents the need for manual annotation, there is currently a large performance gap between methods trained on real and synthetic data. This paper introduces a new method, which bridges this gap.
Most methods trained on synthetic data use 2D images, as domain randomization in 2D is more developed. To obtain precise poses, many of these methods perform a final refinement using 3D data. Our method integrates the 3D data into the network to increase the accuracy of the pose estimation. To allow for domain randomization in 3D, a sensor-based data augmentation has been developed. Additionally, we introduce the SparseEdge feature, which uses a wider search space during point cloud propagation to avoid relying on specific features without increasing run-time.
Experiments on three large pose estimation benchmarks show that the presented method outperforms previous methods trained on synthetic data and achieves comparable results to existing methods trained on real data.

Comments:	10 pages, 5 figures, 7 tables
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:2011.08517 [cs.CV]
	(or arXiv:2011.08517v3 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2011.08517

Submission history

From: Frederik Hagelskjaer [view email]
[v1] Tue, 17 Nov 2020 09:12:11 UTC (8,926 KB)
[v2] Tue, 18 May 2021 13:59:33 UTC (10,865 KB)
[v3] Tue, 17 Aug 2021 09:50:01 UTC (10,860 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Bridging the Reality Gap for Pose Estimation Networks using Sensor-Based Domain Randomization

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Bridging the Reality Gap for Pose Estimation Networks using Sensor-Based Domain Randomization

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators