English

Computer says 'no': Exploring systemic bias in ChatGPT using an audit approach

General Economics 2024-02-13 v3 Economics

Abstract

Large language models offer significant potential for increasing labour productivity, such as streamlining personnel selection, but raise concerns about perpetuating systemic biases embedded into their pre-training data. This study explores the potential ethnic and gender bias of ChatGPT, a chatbot producing human-like responses to language tasks, in assessing job applicants. Using the correspondence audit approach from the social sciences, I simulated a CV screening task with 34,560 vacancy-CV combinations where the chatbot had to rate fictitious applicant profiles. Comparing ChatGPT's ratings of Arab, Asian, Black American, Central African, Dutch, Eastern European, Hispanic, Turkish, and White American male and female applicants, I show that ethnic and gender identity influence the chatbot's evaluations. Ethnic discrimination is more pronounced than gender discrimination and mainly occurs in jobs with favourable labour conditions or requiring greater language proficiency. In contrast, gender discrimination emerges in gender-atypical roles. These findings suggest that ChatGPT's discriminatory output reflects a statistical mechanism echoing societal stereotypes. Policymakers and developers should address systemic bias in language model-driven applications to ensure equitable treatment across demographic groups. Practitioners should practice caution, given the adverse impact these tools can (re)produce, especially in selection decisions involving humans.

Keywords

Cite

@article{arxiv.2309.07664,
  title  = {Computer says 'no': Exploring systemic bias in ChatGPT using an audit approach},
  author = {Louis Lippens},
  journal= {arXiv preprint arXiv:2309.07664},
  year   = {2024}
}

Comments

39 pages, 2 tables, 4 figures; for data and supplementary tables, see https://osf.io/vezt7

R2 v1 2026-06-28T12:21:29.207Z