Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization

Mukherjee, Subhojyoti; Lalitha, Anusha; Sengupta, Sailik; Deshmukh, Aniket; Kveton, Branislav

Computer Science > Machine Learning

arXiv:2412.05469 (cs)

[Submitted on 6 Dec 2024]

Title:Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization

Authors:Subhojyoti Mukherjee, Anusha Lalitha, Sailik Sengupta, Aniket Deshmukh, Branislav Kveton

View PDF HTML (experimental)

Abstract:Multi-objective alignment from human feedback (MOAHF) in large language models (LLMs) is a challenging problem as human preferences are complex, multifaceted, and often conflicting. Recent works on MOAHF considered a-priori multi-objective optimization (MOO), where human preferences are known at training or inference time. In contrast, when human preferences are unknown or difficult to quantify, a natural approach is to cover the Pareto front by multiple diverse solutions. We propose an algorithm HaM for learning diverse LLM policies that maximizes their hypervolume. This is the first application of a-posteriori MOO to MOAHF. HaM is computationally and space efficient, and empirically superior across objectives such as harmlessness, helpfulness, humor, faithfulness, and hallucination, on various datasets.

Subjects:	Machine Learning (cs.LG)
Cite as:	arXiv:2412.05469 [cs.LG]
	(or arXiv:2412.05469v1 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2412.05469

Submission history

From: Subhojyoti Mukherjee [view email]
[v1] Fri, 6 Dec 2024 23:51:47 UTC (16,133 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.LG

< prev | next >

new | recent | 2024-12

Change to browse by:

References & Citations

export BibTeX citation

Computer Science > Machine Learning

Title:Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Multi-Objective Alignment of Large Language Models Through Hypervolume Maximization

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators