Formal Specification, Assessment, and Enforcement of Fairness for Generative AIs

Cheng, Chih-Hong; Wu, Changshun; Ruess, Harald; Zhao, Xingyu; Bensalem, Saddek

Computer Science > Machine Learning

arXiv:2404.16663 (cs)

[Submitted on 25 Apr 2024 (v1), last revised 6 May 2024 (this version, v3)]

Title:Formal Specification, Assessment, and Enforcement of Fairness for Generative AIs

Authors:Chih-Hong Cheng, Changshun Wu, Harald Ruess, Xingyu Zhao, Saddek Bensalem

View PDF HTML (experimental)

Abstract:Reinforcing or even exacerbating societal biases and inequalities will increase significantly as generative AI increasingly produces useful artifacts, from text to images and beyond, for the real world. We address these issues by formally characterizing the notion of fairness for generative AI as a basis for monitoring and enforcing fairness. We define two levels of fairness using the notion of infinite sequences of abstractions of AI-generated artifacts such as text or images. The first is the fairness demonstrated on the generated sequences, which is evaluated only on the outputs while agnostic to the prompts and models used. The second is the inherent fairness of the generative AI model, which requires that fairness be manifested when input prompts are neutral, that is, they do not explicitly instruct the generative AI to produce a particular type of output. We also study relative intersectional fairness to counteract the combinatorial explosion of fairness when considering multiple categories together with lazy fairness enforcement. Finally, fairness monitoring and enforcement are tested against some current generative AI models.

Subjects:	Machine Learning (cs.LG); Artificial Intelligence (cs.AI); Computers and Society (cs.CY); Logic in Computer Science (cs.LO); Software Engineering (cs.SE)
Cite as:	arXiv:2404.16663 [cs.LG]
	(or arXiv:2404.16663v3 [cs.LG] for this version)
	https://doi.org/10.48550/arXiv.2404.16663

Submission history

From: Chih-Hong Cheng [view email]
[v1] Thu, 25 Apr 2024 15:04:27 UTC (3,283 KB)
[v2] Fri, 26 Apr 2024 09:30:25 UTC (3,282 KB)
[v3] Mon, 6 May 2024 06:50:15 UTC (3,283 KB)

Computer Science > Machine Learning

Title:Formal Specification, Assessment, and Enforcement of Fairness for Generative AIs

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Machine Learning

Title:Formal Specification, Assessment, and Enforcement of Fairness for Generative AIs

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators