English

Unsupervised Discovery of Implicit Gender Bias

Computation and Language 2020-10-07 v2

Abstract

Despite their prevalence in society, social biases are difficult to identify, primarily because human judgements in this domain can be unreliable. We take an unsupervised approach to identifying gender bias against women at a comment level and present a model that can surface text likely to contain bias. Our main challenge is forcing the model to focus on signs of implicit bias, rather than other artifacts in the data. Thus, our methodology involves reducing the influence of confounds through propensity matching and adversarial learning. Our analysis shows how biased comments directed towards female politicians contain mixed criticisms, while comments directed towards other female public figures focus on appearance and sexualization. Ultimately, our work offers a way to capture subtle biases in various domains without relying on subjective human judgements.

Keywords

Cite

@article{arxiv.2004.08361,
  title  = {Unsupervised Discovery of Implicit Gender Bias},
  author = {Anjalie Field and Yulia Tsvetkov},
  journal= {arXiv preprint arXiv:2004.08361},
  year   = {2020}
}

Comments

Accepted to EMNLP 2020

R2 v1 2026-06-23T14:55:34.815Z