Challenge: Stereotypes are known to have harmful effects, making their detection critical . current research focuses on detecting and evaluating stereotypical biases .
Approach: They propose a five-tuple definition and provide precise terminologies disentangling stereotypes, antistereotypes, stereotypical bias, and general bias.
Outcome: The proposed framework disentangles stereotypes, antistereotypes, stereotypical bias, and general bias.

Similar Papers

Rethinking Research on Stereotypes: An Analysis through Social Psychological and Computational Perspectives (2026.findings-acl)

Copied to clipboard

Challenge: Existing research on stereotypical biases ignores literature on them and results in resource wastage.
Approach: They argue that stereotypes are social constructs shaping human perception and behavior that can produce harmful outcomes under specific conditions.
Outcome: The proposed models can inherit and amplify stereotypes under certain conditions.
Stereotype Detection as a Catalyst for Enhanced Bias Detection: A Multi-Task Learning Approach (2025.findings-acl)

Copied to clipboard

Challenge: a new study addresses bias and stereotypes in language models by exploring how learning them together improves performance.
Approach: They propose a dataset for bias and stereotype detection that integrates religion, gender, socio-economic status, race, profession, and others.
Outcome: The proposed dataset compares encoder-only models and fine-tuned decoder- only models . the results show that learning stereotypes together improves bias detection .
Blind Men and the Elephant: Diverse Perspectives on Gender Stereotypes in Benchmark Datasets (2025.emnlp-main)

Copied to clipboard

Challenge: Existing benchmarks for measuring gender stereotypical bias in language models are inconsistencies . lack of explicit standards in data gathering can have detrimental effects on results .
Approach: They propose that currently available benchmarks capture only partial facets of gender stereotypes . they apply a framework from social psychology to balance data across components of gender stereotypes based on stereotypical benchmarks.
Outcome: The proposed framework improves correlation between different benchmarks by using simple balancing techniques.
Understanding and Countering Stereotypes: A Computational Approach to the Stereotype Content Model (2021.acl-long)

Copied to clipboard

Challenge: Stereotypical language expresses widely-held beliefs about different social categories.
Approach: They propose a computational approach to interpreting stereotypes in text through the Stereotype Content Model (SCM), a comprehensive causal theory from social psychology.
Outcome: The proposed model compares favourably with survey-based studies in the psychological literature on stereotypes and shows that it is realistic and effective.
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations (2025.emnlp-main)

Copied to clipboard

Challenge: Recent years have seen unprecedented gains in generative AI models' capabilities across modalitieslanguage, image, audio, and video domains across the globe.
Approach: They propose a framework to operationalize stereotypes in generative AI evaluations using social psychological research and NLP data.
Outcome: The proposed framework identifies key components of stereotypes that are crucial in AI evaluation, including the target group, associated attribute, relationship characteristics, perceiving group, and context.
StereoSet: Measuring stereotypical bias in pretrained language models (2021.acl-long)

Copied to clipboard

Challenge: Existing literature on stereotypical biases in language models is limited . current evaluations focus on measuring bias without considering language modeling ability .
Approach: They propose to measure stereotypical biases in four domains: gender, profession, race, and religion . they compare stereotypical and language modeling ability of popular models like BERT, GPT-2, RoBERTa and XLnet .
Outcome: The proposed model shows strong stereotypical biases in gender, profession, race, and religion domains.
Quantifying Stereotypes in Language (2024.eacl-long)

Copied to clipboard

Challenge: Existing studies define a sentence as stereotypical and anti-stereotypical, but they lack a fine-grained quantification of stereotypes.
Approach: They quantify stereotypes in language by annotating a dataset to quantify stereotype of sentences.
Outcome: The proposed models validate the findings of the current studies.
Reinforcement Guided Multi-Task Learning Framework for Low-Resource Stereotype Detection (2022.acl-long)

Copied to clipboard

Challenge: Existing ‘Stereotype Detection’ datasets adopt a diagnostic approach toward large PLMs.
Approach: They propose a multi-task model that leverages the abundance of data-rich neighboring tasks such as hate speech detection, offensive language detection, misogyny detection, etc., to improve the empirical performance.
Outcome: The proposed model achieves significant gains over baselines on hate speech detection, offensive language detection, misogyny detection, etc.
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes (2024.lrec-main)

Copied to clipboard

Challenge: Gender stereotypes are pervasive beliefs about individuals based on their gender that shape societal attitudes, behaviours, and even opportunities.
Approach: They propose eleven strategies to automatically counteract gender stereotypes by generating gender-based counter-stereotypes from a questionnaire to male and female participants.
Outcome: The proposed strategies were perceived as offensive and/or implausible by the raters . humour, perspective-taking, counter-examples, and empathy for the speaker were perceived to be less effective.
Analyzing Stereotypes in Generative Text Inference Tasks (2021.findings-acl)

Copied to clipboard

Challenge: Social psychology studies how social stereotypes are shared as part of cultural knowledge .
Approach: They study how stereotypes manifest when potential targets are situated in neutral contexts . they collect human judgments on the presence of stereotypes in generated inferences based on annotator positionality .
Outcome: The results show that the annotators' positions differ depending on the type of inferences they generate .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations