Theory-Grounded Measurement of U.S. Social Stereotypes in English Language Models

42Citations
Citations of this article
55Readers
Mendeley users who have this article in their library.

Abstract

NLP models trained on text have been shown to reproduce human stereotypes, which can magnify harms to marginalized groups when systems are deployed at scale. We adapt the Agency-Belief-Communion (ABC) stereotype model of Koch et al. (2016) from social psychology as a framework for the systematic study and discovery of stereotypic group-trait associations in language models (LMs). We introduce the sensitivity test (SeT) for measuring stereotypical associations from language models. To evaluate SeT and other measures using the ABC model, we collect group-trait judgments from U.S.-based subjects to compare with English LM stereotypes. Finally, we extend this framework to measure LM stereotyping of intersectional identities.

Cite

CITATION STYLE

APA

Cao, Y. T., Sotnikova, A., Daumé, H., Rudinger, R., & Zou, L. (2022). Theory-Grounded Measurement of U.S. Social Stereotypes in English Language Models. In NAACL 2022 - 2022 Conference of the North American Chapter of the Association for Computational Linguistics: Human Language Technologies, Proceedings of the Conference (pp. 1276–1295). Association for Computational Linguistics (ACL). https://doi.org/10.18653/v1/2022.naacl-main.92

Register to see more suggestions

Mendeley helps you to discover research relevant for your work.

Already have an account?

Save time finding and organizing research with Mendeley

Sign up for free