Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes (2024.lrec-main)
Copied to clipboard
| Challenge: | Gender stereotypes are pervasive beliefs about individuals based on their gender that shape societal attitudes, behaviours, and even opportunities. |
| Approach: | They propose eleven strategies to automatically counteract gender stereotypes by generating gender-based counter-stereotypes from a questionnaire to male and female participants. |
| Outcome: | The proposed strategies were perceived as offensive and/or implausible by the raters . humour, perspective-taking, counter-examples, and empathy for the speaker were perceived to be less effective. |
Similar Papers
Rethinking Research on Stereotypes: An Analysis through Social Psychological and Computational Perspectives (2026.findings-acl)
Copied to clipboard
| Challenge: | Existing research on stereotypical biases ignores literature on them and results in resource wastage. |
| Approach: | They argue that stereotypes are social constructs shaping human perception and behavior that can produce harmful outcomes under specific conditions. |
| Outcome: | The proposed models can inherit and amplify stereotypes under certain conditions. |
Counterfactual Data Augmentation for Mitigating Gender Stereotypes in Languages with Rich Morphology (P19-1)
Copied to clipboard
| Challenge: | Gender stereotypes are manifest in most of the world's languages and are consequently propagated or amplified by NLP systems. |
| Approach: | They propose a method for converting between masculine-inflected and feminine-infflectes sentences in morphologically rich languages to reduce gender stereotyping by a factor of 2.5 without any sacrifice to grammaticality. |
| Outcome: | The proposed approach reduces gender stereotyping by 2.5 without any sacrifice to grammaticality. |
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations (2025.emnlp-main)
Copied to clipboard
| Challenge: | Recent years have seen unprecedented gains in generative AI models' capabilities across modalitieslanguage, image, audio, and video domains across the globe. |
| Approach: | They propose a framework to operationalize stereotypes in generative AI evaluations using social psychological research and NLP data. |
| Outcome: | The proposed framework identifies key components of stereotypes that are crucial in AI evaluation, including the target group, associated attribute, relationship characteristics, perceiving group, and context. |
StereoDetect: Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings (2025.findings-emnlp)
Copied to clipboard
| Challenge: | Stereotypes are known to have harmful effects, making their detection critical . current research focuses on detecting and evaluating stereotypical biases . |
| Approach: | They propose a five-tuple definition and provide precise terminologies disentangling stereotypes, antistereotypes, stereotypical bias, and general bias. |
| Outcome: | The proposed framework disentangles stereotypes, antistereotypes, stereotypical bias, and general bias. |
Understanding and Countering Stereotypes: A Computational Approach to the Stereotype Content Model (2021.acl-long)
Copied to clipboard
| Challenge: | Stereotypical language expresses widely-held beliefs about different social categories. |
| Approach: | They propose a computational approach to interpreting stereotypes in text through the Stereotype Content Model (SCM), a comprehensive causal theory from social psychology. |
| Outcome: | The proposed model compares favourably with survey-based studies in the psychological literature on stereotypes and shows that it is realistic and effective. |
Blind Men and the Elephant: Diverse Perspectives on Gender Stereotypes in Benchmark Datasets (2025.emnlp-main)
Copied to clipboard
| Challenge: | Existing benchmarks for measuring gender stereotypical bias in language models are inconsistencies . lack of explicit standards in data gathering can have detrimental effects on results . |
| Approach: | They propose that currently available benchmarks capture only partial facets of gender stereotypes . they apply a framework from social psychology to balance data across components of gender stereotypes based on stereotypical benchmarks. |
| Outcome: | The proposed framework improves correlation between different benchmarks by using simple balancing techniques. |
Beyond Denouncing Hate: Strategies for Countering Implied Biases and Stereotypes in Language (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Counterspeech, i.e. responses to counteract potential harms of hateful speech, has become an increasingly popular solution to address online hate speech without censorship risks of deletion-based content moderation. |
| Approach: | They draw from psychology and philosophy literature to craft six psychologically inspired strategies to challenge the underlying stereotypical implications of hateful language. |
| Outcome: | The strategies used in human- and machine-generated counterspeech datasets are convincing, whereas human-written counterspech uses less specific strategies compared to machine-produced counters. |
Mitigating Gender Bias in Natural Language Processing: Literature Review (P19-1)
Copied to clipboard
Tony Sun, Andrew Gaut, Shirlyn Tang, Yuxin Huang, Mai ElSherief, Jieyu Zhao, Diba Mirza, Elizabeth Belding, Kai-Wei Chang, William Yang Wang
| Challenge: | NLP models propagate and may even amplify gender bias found in text corpora . methods to mitigate gender bias in NLP are relatively nascent . |
| Approach: | They propose to analyze gender bias based on four forms of representation bias and discuss the advantages and drawbacks of existing gender debiasing methods. |
| Outcome: | The proposed methods are based on four forms of representation bias and have advantages and drawbacks. |
Intersectional Stereotypes in Large Language Models: Dataset and Analysis (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Existing studies on intersectional stereotypes focus on broader, individual categories . current studies focus on single-group stereotypes, such as racial bias against African Americans . |
| Approach: | They propose to use a dataset of intersectional stereotypes curated with the ChatGPT model to analyze propagation in three contemporary LLMs. |
| Outcome: | The proposed dataset enables analysis of stereotype propagation in three contemporary LLMs. |
Exploring Human Gender Stereotypes with Word Association Test (D19-1)
Copied to clipboard
| Challenge: | Existing word embeddings have been used to study gender stereotypes in texts . however, evaluating their validities is still an open problem . et al.: this study investigates gender bias using the lens of language, especially, the words . |
| Approach: | They use word association test to derive bias scores for large amount of words . they find that these bias scores correlate well with bias in the real world . |
| Outcome: | The proposed method correlates well with bias in the real world, and with census data, it provides a different perspective on gender stereotypes in words. |