Benchmarking the rationality of AI decision making using the transitivity axiom

Song, Kiwon; Jennings III, James M.; Davis-Stober, Clintin P.

Computer Science > Artificial Intelligence

arXiv:2502.10554 (cs)

[Submitted on 14 Feb 2025]

Title:Benchmarking the rationality of AI decision making using the transitivity axiom

Authors:Kiwon Song, James M. Jennings III, Clintin P. Davis-Stober

View PDF HTML (experimental)

Abstract:Fundamental choice axioms, such as transitivity of preference, provide testable conditions for determining whether human decision making is rational, i.e., consistent with a utility representation. Recent work has demonstrated that AI systems trained on human data can exhibit similar reasoning biases as humans and that AI can, in turn, bias human judgments through AI recommendation systems. We evaluate the rationality of AI responses via a series of choice experiments designed to evaluate transitivity of preference in humans. We considered ten versions of Meta's Llama 2 and 3 LLM models. We applied Bayesian model selection to evaluate whether these AI-generated choices violated two prominent models of transitivity. We found that the Llama 2 and 3 models generally satisfied transitivity, but when violations did occur, occurred only in the Chat/Instruct versions of the LLMs. We argue that rationality axioms, such as transitivity of preference, can be useful for evaluating and benchmarking the quality of AI-generated responses and provide a foundation for understanding computational rationality in AI systems more generally.

Comments:	13 pages, 2 figures, 3 tables
Subjects:	Artificial Intelligence (cs.AI)
Cite as:	arXiv:2502.10554 [cs.AI]
	(or arXiv:2502.10554v1 [cs.AI] for this version)
	https://doi.org/10.48550/arXiv.2502.10554

Submission history

From: Clintin Davis-Stober [view email]
[v1] Fri, 14 Feb 2025 20:56:40 UTC (85 KB)

Computer Science > Artificial Intelligence

Title:Benchmarking the rationality of AI decision making using the transitivity axiom

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Artificial Intelligence

Title:Benchmarking the rationality of AI decision making using the transitivity axiom

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators