Rethinking Research on Stereotypes: An Analysis through Social Psychological and Computational Perspectives (2026.findings-acl)
Copied to clipboard
| Challenge: | Existing research on stereotypical biases ignores literature on them and results in resource wastage. |
| Approach: | They argue that stereotypes are social constructs shaping human perception and behavior that can produce harmful outcomes under specific conditions. |
| Outcome: | The proposed models can inherit and amplify stereotypes under certain conditions. |
Similar Papers
A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations (2025.emnlp-main)
Copied to clipboard
| Challenge: | Recent years have seen unprecedented gains in generative AI models' capabilities across modalitieslanguage, image, audio, and video domains across the globe. |
| Approach: | They propose a framework to operationalize stereotypes in generative AI evaluations using social psychological research and NLP data. |
| Outcome: | The proposed framework identifies key components of stereotypes that are crucial in AI evaluation, including the target group, associated attribute, relationship characteristics, perceiving group, and context. |
Understanding and Countering Stereotypes: A Computational Approach to the Stereotype Content Model (2021.acl-long)
Copied to clipboard
| Challenge: | Stereotypical language expresses widely-held beliefs about different social categories. |
| Approach: | They propose a computational approach to interpreting stereotypes in text through the Stereotype Content Model (SCM), a comprehensive causal theory from social psychology. |
| Outcome: | The proposed model compares favourably with survey-based studies in the psychological literature on stereotypes and shows that it is realistic and effective. |
StereoDetect: Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings (2025.findings-emnlp)
Copied to clipboard
| Challenge: | Stereotypes are known to have harmful effects, making their detection critical . current research focuses on detecting and evaluating stereotypical biases . |
| Approach: | They propose a five-tuple definition and provide precise terminologies disentangling stereotypes, antistereotypes, stereotypical bias, and general bias. |
| Outcome: | The proposed framework disentangles stereotypes, antistereotypes, stereotypical bias, and general bias. |
Quantifying Stereotypes in Language (2024.eacl-long)
Copied to clipboard
| Challenge: | Existing studies define a sentence as stereotypical and anti-stereotypical, but they lack a fine-grained quantification of stereotypes. |
| Approach: | They quantify stereotypes in language by annotating a dataset to quantify stereotype of sentences. |
| Outcome: | The proposed models validate the findings of the current studies. |
Intersectional Stereotypes in Large Language Models: Dataset and Analysis (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Existing studies on intersectional stereotypes focus on broader, individual categories . current studies focus on single-group stereotypes, such as racial bias against African Americans . |
| Approach: | They propose to use a dataset of intersectional stereotypes curated with the ChatGPT model to analyze propagation in three contemporary LLMs. |
| Outcome: | The proposed dataset enables analysis of stereotype propagation in three contemporary LLMs. |
Who is better at math, Jenny or Jingzhen? Uncovering Stereotypes in Large Language Models (2024.emnlp-main)
Copied to clipboard
| Challenge: | Existing research on stereotypes in large language models is limited and focuses on African Ameri- F. |
| Approach: | They propose to use global bias to probe a set of large language models via perplexity to determine how certain stereotypes are represented in the model's internal representations. |
| Outcome: | The proposed model amplifys harmful stereotypes and shows that the demographic groups associated with stereotypes remain consistent across model likelihoods and outputs. |
Uncovering Stereotypes in Large Language Models: A Task Complexity-based Approach (2024.eacl-long)
Copied to clipboard
| Challenge: | Recent Large Language Models (LLMs) have unlocked unprecedented applications of AI. |
| Approach: | They propose to use a social benchmark to evaluate the bias protection provided by Large Language Models (LLMs) with a variety of tasks with varying complexities to assess their effectiveness. |
| Outcome: | The proposed benchmark shows that both ChatGPT and GPT-4 have strong biases with respect to nationality, gender, race, and religion. |
StereoMap: Quantifying the Awareness of Human-like Stereotypes in Large Language Models (2023.emnlp-main)
Copied to clipboard
| Challenge: | Large Language Models (LLMs) have been observed to encode harmful associations present in the training data. |
| Approach: | They propose a framework to map LLMs' perceptions of how demographic groups have been viewed by society using the dimensions of Warmth and Competence. |
| Outcome: | The proposed framework maps LLMs’ perceptions of social groups using the dimensions of Warmth and Competence. |
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes (2024.lrec-main)
Copied to clipboard
| Challenge: | Gender stereotypes are pervasive beliefs about individuals based on their gender that shape societal attitudes, behaviours, and even opportunities. |
| Approach: | They propose eleven strategies to automatically counteract gender stereotypes by generating gender-based counter-stereotypes from a questionnaire to male and female participants. |
| Outcome: | The proposed strategies were perceived as offensive and/or implausible by the raters . humour, perspective-taking, counter-examples, and empathy for the speaker were perceived to be less effective. |
Auditing LLM Responses to Harmful Stereotypes Targeting Mental Health Groups (2026.findings-acl)
Copied to clipboard
| Challenge: | Large Language Models (LLMs) can exhibit imbalanced biases against vulnerable groups, but how they rationalize stereotypes and rights restrictions targeting mental health entities remains underexplored. |
| Approach: | They audit a suite of open-weight LLMs on stereotype-justification prompts tied to mental health identities. |
| Outcome: | The proposed models endorse harmful stereotypes when explicitly asked to justify them, with endorsement varying across model families, versions, and mental health conditions. |