arXiv Open Access 2022

Challenges in Measuring Bias via Open-Ended Language Generation

Afra Feyza Akyürek Muhammed Yusuf Kocyigit Sejin Paik Derry Wijaya

Lihat Sumber

Abstrak

Researchers have devised numerous ways to quantify social biases vested in pretrained language models. As some language models are capable of generating coherent completions given a set of textual prompts, several prompting datasets have been proposed to measure biases between social groups -- posing language generation as a way of identifying biases. In this opinion paper, we analyze how specific choices of prompt sets, metrics, automatic tools and sampling strategies affect bias results. We find out that the practice of measuring biases through text completion is prone to yielding contradicting results under different experiment settings. We additionally provide recommendations for reporting biases in open-ended language generation for a more complete outlook of biases exhibited by a given language model. Code to reproduce the results is released under https://github.com/feyzaakyurek/bias-textgen.

Topik & Kata Kunci

cs.CL cs.CY

Penulis (4)

Afra Feyza Akyürek

Muhammed Yusuf Kocyigit

Sejin Paik

Derry Wijaya

Format Sitasi

APA MLA BibTeX

Akyürek, A.F., Kocyigit, M.Y., Paik, S., Wijaya, D. (2022). Challenges in Measuring Bias via Open-Ended Language Generation. https://arxiv.org/abs/2205.11601

Akses Cepat

Lihat di Sumber

Informasi Jurnal

Tahun Terbit: 2022
Bahasa: en
Sumber Database: arXiv
Akses: Open Access ✓