RankAlign: A Ranking View of the Generator-Validator Gap in Large Language Models

Rodriguez, Juan Diego; Ding, Wenxuan; Erk, Katrin; Durrett, Greg

Computer Science > Computation and Language

arXiv:2504.11381 (cs)

[Submitted on 15 Apr 2025]

Title:RankAlign: A Ranking View of the Generator-Validator Gap in Large Language Models

Authors:Juan Diego Rodriguez, Wenxuan Ding, Katrin Erk, Greg Durrett

View PDF HTML (experimental)

Abstract:Although large language models (LLMs) have become generally more capable and accurate across many tasks, some fundamental sources of unreliability remain in their behavior. One key limitation is their inconsistency at reporting the the same information when prompts are changed. In this paper, we consider the discrepancy between a model's generated answer and their own verification of that answer, the generator-validator gap. We define this gap in a more stringent way than prior work: we expect correlation of scores from a generator and a validator over the entire set of candidate answers. We show that according to this measure, a large gap exists in various settings, including question answering, lexical semantics tasks, and next-word prediction. We then propose RankAlign, a ranking-based training method, and show that it significantly closes the gap by 31.8% on average, surpassing all baseline methods. Moreover, this approach generalizes well to out-of-domain tasks and lexical items.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2504.11381 [cs.CL]
	(or arXiv:2504.11381v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2504.11381

Submission history

From: Juan Diego Rodriguez [view email]
[v1] Tue, 15 Apr 2025 16:53:31 UTC (6,387 KB)

Computer Science > Computation and Language

Title:RankAlign: A Ranking View of the Generator-Validator Gap in Large Language Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:RankAlign: A Ranking View of the Generator-Validator Gap in Large Language Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators