Semantic Properties of cosine based bias scores for word embeddings

Schröder, Sarah; Schulz, Alexander; Hinder, Fabian; Hammer, Barbara

Computer Science > Computation and Language

arXiv:2401.15499 (cs)

[Submitted on 27 Jan 2024]

Title:Semantic Properties of cosine based bias scores for word embeddings

Authors:Sarah Schröder, Alexander Schulz, Fabian Hinder, Barbara Hammer

View PDF

Abstract:Plenty of works have brought social biases in language models to attention and proposed methods to detect such biases. As a result, the literature contains a great deal of different bias tests and scores, each introduced with the premise to uncover yet more biases that other scores fail to detect. What severely lacks in the literature, however, are comparative studies that analyse such bias scores and help researchers to understand the benefits or limitations of the existing methods. In this work, we aim to close this gap for cosine based bias scores. By building on a geometric definition of bias, we propose requirements for bias scores to be considered meaningful for quantifying biases. Furthermore, we formally analyze cosine based scores from the literature with regard to these requirements. We underline these findings with experiments to show that the bias scores' limitations have an impact in the application case.

Comments:	11 pages, 4 figures
Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2401.15499 [cs.CL]
	(or arXiv:2401.15499v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2401.15499

Submission history

From: Sarah Schröder [view email]
[v1] Sat, 27 Jan 2024 20:31:10 UTC (146 KB)

Computer Science > Computation and Language

Title:Semantic Properties of cosine based bias scores for word embeddings

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Semantic Properties of cosine based bias scores for word embeddings

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators