Theory-Grounded Measurement of U.S. Social Stereotypes in English Language Models

Cao, Yang Trista; Sotnikova, Anna; Daumé III, Hal; Rudinger, Rachel; Zou, Linda

Computer Science > Computation and Language

arXiv:2206.11684 (cs)

[Submitted on 23 Jun 2022]

Title:Theory-Grounded Measurement of U.S. Social Stereotypes in English Language Models

Authors:Yang Trista Cao, Anna Sotnikova, Hal Daumé III, Rachel Rudinger, Linda Zou

View PDF

Abstract:NLP models trained on text have been shown to reproduce human stereotypes, which can magnify harms to marginalized groups when systems are deployed at scale. We adapt the Agency-Belief-Communion (ABC) stereotype model of Koch et al. (2016) from social psychology as a framework for the systematic study and discovery of stereotypic group-trait associations in language models (LMs). We introduce the sensitivity test (SeT) for measuring stereotypical associations from language models. To evaluate SeT and other measures using the ABC model, we collect group-trait judgments from U.S.-based subjects to compare with English LM stereotypes. Finally, we extend this framework to measure LM stereotyping of intersectional identities.

Subjects:	Computation and Language (cs.CL)
Cite as:	arXiv:2206.11684 [cs.CL]
	(or arXiv:2206.11684v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2206.11684

Submission history

From: Anna Sotnikova [view email]
[v1] Thu, 23 Jun 2022 13:22:24 UTC (844 KB)

Full-text links:

Access Paper:

view license

Current browse context:

cs.CL

< prev | next >

new | recent | 2022-06

Change to browse by:

References & Citations

export BibTeX citation

Computer Science > Computation and Language

Title:Theory-Grounded Measurement of U.S. Social Stereotypes in English Language Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computation and Language

Title:Theory-Grounded Measurement of U.S. Social Stereotypes in English Language Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators