PaperScope
LIVE · 2026-09-03 05:40 UTC

Evaluating and Mitigating Anti-LGBTQ Biases in German and Multilingual Language Models

Melina Morch, Daniel Braun

Latestcs.CLcs.LGcs.AIcs.CV
arXiv ID
2608.30884 v1
Category
Submitted
2026-08-31

Abstract

While gender and racial biases in language models have been widely studied, anti-LGBTQ biases remain underexplored, particularly beyond English. Existing benchmarks often do not capture cultural and linguistic variation and rely on gender representations. This paper introduces a multilingual German-English benchmark dataset for the evaluation of anti-LGBTQ biases in language models. It combines community-sourced stereotypes from German-speaking queer individuals with a German translation of WinoQueer. The data is used to evaluate eight language models across sizes and architectures and explore mitigation through fine-tuning on community and progressive media content. Results show that language models reproduce anti-queer stereotypes, with variation across identities and models. Differences between the translated and community-based data highlight the importance of cultural adaptation for multilingual bias evaluation. Fine-tuning reduces bias on average, but not consistently across models and identities. Warning: This text contains examples of anti-queer hateful language and stereotypes.

Comment: Accepted at EMNLP 2026

arXiv abs page · PDF · same-day batch