SpaRC: Sparse Radar-Camera Fusion for 3D Object Detection

Wolters, Philipp; Gilg, Johannes; Teepe, Torben; Herzog, Fabian; Fent, Felix; Rigoll, Gerhard

Computer Science > Computer Vision and Pattern Recognition

arXiv:2411.19860 (cs)

[Submitted on 29 Nov 2024]

Title:SpaRC: Sparse Radar-Camera Fusion for 3D Object Detection

Authors:Philipp Wolters, Johannes Gilg, Torben Teepe, Fabian Herzog, Felix Fent, Gerhard Rigoll

View PDF HTML (experimental)

Abstract:In this work, we present SpaRC, a novel Sparse fusion transformer for 3D perception that integrates multi-view image semantics with Radar and Camera point features. The fusion of radar and camera modalities has emerged as an efficient perception paradigm for autonomous driving systems. While conventional approaches utilize dense Bird's Eye View (BEV)-based architectures for depth estimation, contemporary query-based transformers excel in camera-only detection through object-centric methodology. However, these query-based approaches exhibit limitations in false positive detections and localization precision due to implicit depth modeling. We address these challenges through three key contributions: (1) sparse frustum fusion (SFF) for cross-modal feature alignment, (2) range-adaptive radar aggregation (RAR) for precise object localization, and (3) local self-attention (LSA) for focused query aggregation. In contrast to existing methods requiring computationally intensive BEV-grid rendering, SpaRC operates directly on encoded point features, yielding substantial improvements in efficiency and accuracy. Empirical evaluations on the nuScenes and TruckScenes benchmarks demonstrate that SpaRC significantly outperforms existing dense BEV-based and sparse query-based detectors. Our method achieves state-of-the-art performance metrics of 67.1 NDS and 63.1 AMOTA. The code and pretrained models are available at this https URL.

Comments:	18 pages, 11 figures
Subjects:	Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
Cite as:	arXiv:2411.19860 [cs.CV]
	(or arXiv:2411.19860v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2411.19860

Submission history

From: Philipp Wolters [view email]
[v1] Fri, 29 Nov 2024 17:17:38 UTC (22,961 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:SpaRC: Sparse Radar-Camera Fusion for 3D Object Detection

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:SpaRC: Sparse Radar-Camera Fusion for 3D Object Detection

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators