Debiasing Sentence Embedders through Contrastive Word Pairs

Kenneweg, Philip; Schröder, Sarah; Schulz, Alexander; Hammer, Barbara

doi:10.5220/0011615300003411

Computer Science > Computation and Language

arXiv:2403.18555 (cs)

[Submitted on 27 Mar 2024]

Title:Debiasing Sentence Embedders through Contrastive Word Pairs

Authors:Philip Kenneweg, Sarah Schröder, Alexander Schulz, Barbara Hammer

View PDF HTML (experimental)

Abstract:Over the last years, various sentence embedders have been an integral part in the success of current machine learning approaches to Natural Language Processing (NLP). Unfortunately, multiple sources have shown that the bias, inherent in the datasets upon which these embedding methods are trained, is learned by them. A variety of different approaches to remove biases in embeddings exists in the literature. Most of these approaches are applicable to word embeddings and in fewer cases to sentence embeddings. It is problematic that most debiasing approaches are directly transferred from word embeddings, therefore these approaches fail to take into account the nonlinear nature of sentence embedders and the embeddings they produce. It has been shown in literature that bias information is still present if sentence embeddings are debiased using such methods. In this contribution, we explore an approach to remove linear and nonlinear bias information for NLP solutions, without impacting downstream performance. We compare our approach to common debiasing methods on classical bias metrics and on bias metrics which take nonlinear information into account.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2403.18555 [cs.CL]
	(or arXiv:2403.18555v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2403.18555
Related DOI:	https://doi.org/10.5220/0011615300003411

Submission history

From: Philip Kenneweg [view email]
[v1] Wed, 27 Mar 2024 13:34:59 UTC (926 KB)

Computer Science > Computation and Language

Title:Debiasing Sentence Embedders through Contrastive Word Pairs

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Debiasing Sentence Embedders through Contrastive Word Pairs

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators