LEGO: Learning Edge with Geometry all at Once by Watching Videos

Yang, Zhenheng; Wang, Peng; Wang, Yang; Xu, Wei; Nevatia, Ram

Computer Science > Computer Vision and Pattern Recognition

arXiv:1803.05648v1 (cs)

[Submitted on 15 Mar 2018 (this version), latest version 24 Mar 2018 (v2)]

Title:LEGO: Learning Edge with Geometry all at Once by Watching Videos

Authors:Zhenheng Yang, Peng Wang, Yang Wang, Wei Xu, Ram Nevatia

View PDF

Abstract:Learning to estimate 3D geometry in a single image by watching unlabeled videos via deep convolutional network is attracting significant attention. In this paper, we introduce a "3D as-smooth-as-possible (3D-ASAP)" priori inside the pipeline, which enables joint estimation of edges and 3D scene, yielding results with significant improvement in accuracy for fine detailed structures. Specifically, we define the 3D-ASAP priori by requiring that any two points recovered in 3D from an image should lie on an existing planar surface if no other cues provided. We design an unsupervised framework that Learns Edges and Geometry (depth, normal) all at Once (LEGO). The predicted edges are embedded into depth and surface normal smoothness terms, where pixels without edges in-between are constrained to satisfy the priori. In our framework, the predicted depths, normals and edges are forced to be consistent all the time. We conduct experiments on KITTI to evaluate our estimated geometry and CityScapes to perform edge evaluation. We show that in all of the tasks, this http URL, normal and edge, our algorithm vastly outperforms other state-of-the-art (SOTA) algorithms, demonstrating the benefits of our approach.

Comments:	Accepted to CVPR 2018 as spotlight
Subjects:	Computer Vision and Pattern Recognition (cs.CV)
Cite as:	arXiv:1803.05648 [cs.CV]
	(or arXiv:1803.05648v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.1803.05648

Submission history

From: Zhenheng Yang [view email]
[v1] Thu, 15 Mar 2018 09:14:11 UTC (6,569 KB)
[v2] Sat, 24 Mar 2018 00:31:11 UTC (8,670 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:LEGO: Learning Edge with Geometry all at Once by Watching Videos

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:LEGO: Learning Edge with Geometry all at Once by Watching Videos

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators