Challenge: Existing research on stereotypical biases ignores literature on them and results in resource wastage.
Approach: They argue that stereotypes are social constructs shaping human perception and behavior that can produce harmful outcomes under specific conditions.
Outcome: The proposed models can inherit and amplify stereotypes under certain conditions.

Similar Papers

A Comprehensive Framework to Operationalize Social Stereotypes for Responsible AI Evaluations (2025.emnlp-main)

Copied to clipboard

Challenge: Recent years have seen unprecedented gains in generative AI models' capabilities across modalitieslanguage, image, audio, and video domains across the globe.
Approach: They propose a framework to operationalize stereotypes in generative AI evaluations using social psychological research and NLP data.
Outcome: The proposed framework identifies key components of stereotypes that are crucial in AI evaluation, including the target group, associated attribute, relationship characteristics, perceiving group, and context.
Understanding and Countering Stereotypes: A Computational Approach to the Stereotype Content Model (2021.acl-long)

Copied to clipboard

Challenge: Stereotypical language expresses widely-held beliefs about different social categories.
Approach: They propose a computational approach to interpreting stereotypes in text through the Stereotype Content Model (SCM), a comprehensive causal theory from social psychology.
Outcome: The proposed model compares favourably with survey-based studies in the psychological literature on stereotypes and shows that it is realistic and effective.
StereoDetect: Detecting Stereotypes and Anti-stereotypes the Correct Way Using Social Psychological Underpinnings (2025.findings-emnlp)

Copied to clipboard

Challenge: Stereotypes are known to have harmful effects, making their detection critical . current research focuses on detecting and evaluating stereotypical biases .
Approach: They propose a five-tuple definition and provide precise terminologies disentangling stereotypes, antistereotypes, stereotypical bias, and general bias.
Outcome: The proposed framework disentangles stereotypes, antistereotypes, stereotypical bias, and general bias.
Quantifying Stereotypes in Language (2024.eacl-long)

Copied to clipboard

Challenge: Existing studies define a sentence as stereotypical and anti-stereotypical, but they lack a fine-grained quantification of stereotypes.
Approach: They quantify stereotypes in language by annotating a dataset to quantify stereotype of sentences.
Outcome: The proposed models validate the findings of the current studies.
Intersectional Stereotypes in Large Language Models: Dataset and Analysis (2023.findings-emnlp)

Copied to clipboard

Challenge: Existing studies on intersectional stereotypes focus on broader, individual categories . current studies focus on single-group stereotypes, such as racial bias against African Americans .
Approach: They propose to use a dataset of intersectional stereotypes curated with the ChatGPT model to analyze propagation in three contemporary LLMs.
Outcome: The proposed dataset enables analysis of stereotype propagation in three contemporary LLMs.
Who is better at math, Jenny or Jingzhen? Uncovering Stereotypes in Large Language Models (2024.emnlp-main)

Copied to clipboard

Challenge: Existing research on stereotypes in large language models is limited and focuses on African Ameri- F.
Approach: They propose to use global bias to probe a set of large language models via perplexity to determine how certain stereotypes are represented in the model's internal representations.
Outcome: The proposed model amplifys harmful stereotypes and shows that the demographic groups associated with stereotypes remain consistent across model likelihoods and outputs.
Uncovering Stereotypes in Large Language Models: A Task Complexity-based Approach (2024.eacl-long)

Copied to clipboard

Challenge: Recent Large Language Models (LLMs) have unlocked unprecedented applications of AI.
Approach: They propose to use a social benchmark to evaluate the bias protection provided by Large Language Models (LLMs) with a variety of tasks with varying complexities to assess their effectiveness.
Outcome: The proposed benchmark shows that both ChatGPT and GPT-4 have strong biases with respect to nationality, gender, race, and religion.
StereoMap: Quantifying the Awareness of Human-like Stereotypes in Large Language Models (2023.emnlp-main)

Copied to clipboard

Challenge: Large Language Models (LLMs) have been observed to encode harmful associations present in the training data.
Approach: They propose a framework to map LLMs' perceptions of how demographic groups have been viewed by society using the dimensions of Warmth and Competence.
Outcome: The proposed framework maps LLMs’ perceptions of social groups using the dimensions of Warmth and Competence.
Challenging Negative Gender Stereotypes: A Study on the Effectiveness of Automated Counter-Stereotypes (2024.lrec-main)

Copied to clipboard

Challenge: Gender stereotypes are pervasive beliefs about individuals based on their gender that shape societal attitudes, behaviours, and even opportunities.
Approach: They propose eleven strategies to automatically counteract gender stereotypes by generating gender-based counter-stereotypes from a questionnaire to male and female participants.
Outcome: The proposed strategies were perceived as offensive and/or implausible by the raters . humour, perspective-taking, counter-examples, and empathy for the speaker were perceived to be less effective.
Auditing LLM Responses to Harmful Stereotypes Targeting Mental Health Groups (2026.findings-acl)

Copied to clipboard

Challenge: Large Language Models (LLMs) can exhibit imbalanced biases against vulnerable groups, but how they rationalize stereotypes and rights restrictions targeting mental health entities remains underexplored.
Approach: They audit a suite of open-weight LLMs on stereotype-justification prompts tied to mental health identities.
Outcome: The proposed models endorse harmful stereotypes when explicitly asked to justify them, with endorsement varying across model families, versions, and mental health conditions.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations