To Rely or Not to Rely? Evaluating Interventions for Appropriate Reliance on Large Language Models

Bo, Jessica Y.; Wan, Sophia; Anderson, Ashton

Computer Science > Human-Computer Interaction

arXiv:2412.15584 (cs)

[Submitted on 20 Dec 2024]

Title:To Rely or Not to Rely? Evaluating Interventions for Appropriate Reliance on Large Language Models

Authors:Jessica Y. Bo, Sophia Wan, Ashton Anderson

View PDF HTML (experimental)

Abstract:As Large Language Models become integral to decision-making, optimism about their power is tempered with concern over their errors. Users may over-rely on LLM advice that is confidently stated but wrong, or under-rely due to mistrust. Reliance interventions have been developed to help users of LLMs, but they lack rigorous evaluation for appropriate reliance. We benchmark the performance of three relevant interventions by conducting a randomized online experiment with 400 participants attempting two challenging tasks: LSAT logical reasoning and image-based numerical estimation. For each question, participants first answered independently, then received LLM advice modified by one of three reliance interventions and answered the question again. Our findings indicate that while interventions reduce over-reliance, they generally fail to improve appropriate reliance. Furthermore, people became more confident after making incorrect reliance decisions in certain contexts, demonstrating poor calibration. Based on our findings, we discuss implications for designing effective reliance interventions in human-LLM collaboration.

Subjects:	Human-Computer Interaction (cs.HC)
Cite as:	arXiv:2412.15584 [cs.HC]
	(or arXiv:2412.15584v1 [cs.HC] for this version)
	https://doi.org/10.48550/arXiv.2412.15584

Submission history

From: Jessica Bo [view email]
[v1] Fri, 20 Dec 2024 05:40:32 UTC (4,745 KB)

Computer Science > Human-Computer Interaction

Title:To Rely or Not to Rely? Evaluating Interventions for Appropriate Reliance on Large Language Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Human-Computer Interaction

Title:To Rely or Not to Rely? Evaluating Interventions for Appropriate Reliance on Large Language Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators