Explainable deep learning improves human mental models of self-driving cars

Kenny, Eoin M.; Dharmavaram, Akshay; Lee, Sang Uk; Phan-Minh, Tung; Rajesh, Shreyas; Hu, Yunqing; Major, Laura; Tomov, Momchil S.; Shah, Julie A.

Computer Science > Robotics

arXiv:2411.18714 (cs)

[Submitted on 27 Nov 2024]

Title:Explainable deep learning improves human mental models of self-driving cars

Authors:Eoin M. Kenny, Akshay Dharmavaram, Sang Uk Lee, Tung Phan-Minh, Shreyas Rajesh, Yunqing Hu, Laura Major, Momchil S. Tomov, Julie A. Shah

View PDF HTML (experimental)

Abstract:Self-driving cars increasingly rely on deep neural networks to achieve human-like driving. However, the opacity of such black-box motion planners makes it challenging for the human behind the wheel to accurately anticipate when they will fail, with potentially catastrophic consequences. Here, we introduce concept-wrapper network (i.e., CW-Net), a method for explaining the behavior of black-box motion planners by grounding their reasoning in human-interpretable concepts. We deploy CW-Net on a real self-driving car and show that the resulting explanations refine the human driver's mental model of the car, allowing them to better predict its behavior and adjust their own behavior accordingly. Unlike previous work using toy domains or simulations, our study presents the first real-world demonstration of how to build authentic autonomous vehicles (AVs) that give interpretable, causally faithful explanations for their decisions, without sacrificing performance. We anticipate our method could be applied to other safety-critical systems with a human in the loop, such as autonomous drones and robotic surgeons. Overall, our study suggests a pathway to explainability for autonomous agents as a whole, which can help make them more transparent, their deployment safer, and their usage more ethical.

Comments:	* - equal contribution
Subjects:	Robotics (cs.RO); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2411.18714 [cs.RO]
	(or arXiv:2411.18714v1 [cs.RO] for this version)
	https://doi.org/10.48550/arXiv.2411.18714

Submission history

From: Momchil Tomov [view email]
[v1] Wed, 27 Nov 2024 19:38:43 UTC (7,943 KB)

Computer Science > Robotics

Title:Explainable deep learning improves human mental models of self-driving cars

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Robotics

Title:Explainable deep learning improves human mental models of self-driving cars

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators