CO^3: Cooperative Unsupervised 3D Representation Learning for Autonomous Driving

Chen, Runjian; Mu, Yao; Xu, Runsen; Shao, Wenqi; Jiang, Chenhan; Xu, Hang; Li, Zhenguo; Luo, Ping

Computer Science > Computer Vision and Pattern Recognition

arXiv:2206.04028 (cs)

[Submitted on 8 Jun 2022 (v1), last revised 26 Jun 2022 (this version, v2)]

Title:CO^3: Cooperative Unsupervised 3D Representation Learning for Autonomous Driving

Authors:Runjian Chen, Yao Mu, Runsen Xu, Wenqi Shao, Chenhan Jiang, Hang Xu, Zhenguo Li, Ping Luo

View PDF

Abstract:Unsupervised contrastive learning for indoor-scene point clouds has achieved great successes. However, unsupervised learning point clouds in outdoor scenes remains challenging because previous methods need to reconstruct the whole scene and capture partial views for the contrastive objective. This is infeasible in outdoor scenes with moving objects, obstacles, and sensors. In this paper, we propose CO^3, namely Cooperative Contrastive Learning and Contextual Shape Prediction, to learn 3D representation for outdoor-scene point clouds in an unsupervised manner. CO^3 has several merits compared to existing methods. (1) It utilizes LiDAR point clouds from vehicle-side and infrastructure-side to build views that differ enough but meanwhile maintain common semantic information for contrastive learning, which are more appropriate than views built by previous methods. (2) Alongside the contrastive objective, shape context prediction is proposed as pre-training goal and brings more task-relevant information for unsupervised 3D point cloud representation learning, which are beneficial when transferring the learned representation to downstream detection tasks. (3) As compared to previous methods, representation learned by CO^3 is able to be transferred to different outdoor scene dataset collected by different type of LiDAR sensors. (4) CO^3 improves current state-of-the-art methods on both Once and KITTI datasets by up to 2.58 mAP. Codes and models will be released. We believe CO^3 will facilitate understanding LiDAR point clouds in outdoor scene.

Comments:	Pre-trained backbones and fine-tuned downstream models are now available: this https URL. Code will be released
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Robotics (cs.RO)
Cite as:	arXiv:2206.04028 [cs.CV]
	(or arXiv:2206.04028v2 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2206.04028

Submission history

From: Runjian Chen [view email]
[v1] Wed, 8 Jun 2022 17:37:58 UTC (14,483 KB)
[v2] Sun, 26 Jun 2022 10:36:04 UTC (14,495 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:CO^3: Cooperative Unsupervised 3D Representation Learning for Autonomous Driving

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:CO^3: Cooperative Unsupervised 3D Representation Learning for Autonomous Driving

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators