The Impact of Input Order Bias on Large Language Models for Software Fault Localization

Rafi, Md Nakhla; Kim, Dong Jae; Chen, Tse-Hsun; Wang, Shaowei

Computer Science > Software Engineering

arXiv:2412.18750 (cs)

[Submitted on 25 Dec 2024]

Title:The Impact of Input Order Bias on Large Language Models for Software Fault Localization

Authors:Md Nakhla Rafi, Dong Jae Kim, Tse-Hsun Chen, Shaowei Wang

View PDF

Abstract:Large Language Models (LLMs) show great promise in software engineering tasks like Fault Localization (FL) and Automatic Program Repair (APR). This study examines how input order and context size affect LLM performance in FL, a key step for many downstream software engineering tasks. We test different orders for methods using Kendall Tau distances, including "perfect" (where ground truths come first) and "worst" (where ground truths come last). Our results show a strong bias in order, with Top-1 accuracy falling from 57\% to 20\% when we reverse the code order. Breaking down inputs into smaller contexts helps reduce this bias, narrowing the performance gap between perfect and worst orders from 22\% to just 1\%. We also look at ordering methods based on traditional FL techniques and metrics. Ordering using DepGraph's ranking achieves 48\% Top-1 accuracy, better than more straightforward ordering approaches like CallGraph. These findings underscore the importance of how we structure inputs, manage contexts, and choose ordering methods to improve LLM performance in FL and other software engineering tasks.

Subjects:	Software Engineering (cs.SE); Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
Cite as:	arXiv:2412.18750 [cs.SE]
	(or arXiv:2412.18750v1 [cs.SE] for this version)
	https://doi.org/10.48550/arXiv.2412.18750

Submission history

From: Md Nakhla Rafi [view email]
[v1] Wed, 25 Dec 2024 02:48:53 UTC (1,962 KB)

Computer Science > Software Engineering

Title:The Impact of Input Order Bias on Large Language Models for Software Fault Localization

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Software Engineering

Title:The Impact of Input Order Bias on Large Language Models for Software Fault Localization

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators