Social Biases in NLP Models as Barriers for Persons with Disabilities (2020.acl-main)
Copied to clipboard
| Challenge: | toxicity prediction and sentiment analysis models perpetuate undesirable social biases from the data on which they are trained. |
| Approach: | They propose to use toxicity prediction and sentiment analysis to examine whether NLP models perpetuate undesirable biases towards mentions of disability. |
| Outcome: | The proposed models contain undesirable biases towards mentions of disability in two English language models. |
Similar Papers
Unpacking the Interdependent Systems of Discrimination: Ableist Bias in NLP Systems through an Intersectional Lens (2021.findings-emnlp)
Copied to clipboard
| Challenge: | Statistically significant results demonstrate that people with disabilities can be disadvantaged. |
| Approach: | They used a large-scale BERT language model to predict word predictions and found that people with disabilities can be disadvantaged. |
| Outcome: | The results show that people with disabilities can be disadvantaged and that gender and race identities can be discriminated against. |
A Study of Implicit Bias in Pretrained Language Models against People with Disabilities (2022.coling-1)
Copied to clipboard
| Challenge: | Pretrained language models exhibit sociodemographic biases, such as against gender and race, raising concerns of downstream biase in language technologies. |
| Approach: | They propose to use word embedding-based and transformer-based PLMs to test for the presence of biases against people with disabilities (PWDs) |
| Outcome: | The proposed models favor ableist language, despite their sociodemographic biases against race and gender. |
Bias and Fairness in Natural Language Processing (D19-2)
Copied to clipboard
| Challenge: | a tutorial will review the history of bias and fairness studies in machine learning and language processing . |
| Approach: | This tutorial reviews the history of bias and fairness studies in machine learning and language processing . it presents recent community effort to quantify and mitigat bias in natural language processing models . |
| Outcome: | This tutorial reviews the history of bias and fairness studies in machine learning and language processing . it aims to quantify and mitigate bias in natural language processing models for a wide spectrum of tasks . |
Language (Technology) is Power: A Critical Survey of “Bias” in NLP (2020.acl-main)
Copied to clipboard
| Challenge: | 146 papers analyzing "bias" in NLP systems lack normative reasoning, we find . authors propose three recommendations for work analyzing “bias” in Nlp systems . |
| Approach: | They propose three recommendations for analyzing "bias" in NLP systems . they propose to focus on what kinds of system behaviors are harmful, in what ways, to whom, and why . |
| Outcome: | The proposed methods for measuring or mitigating “bias” are poorly matched to their motivations and do not engage critically with literature outside of NLP. |
Rethinking Research on Stereotypes: An Analysis through Social Psychological and Computational Perspectives (2026.findings-acl)
Copied to clipboard
| Challenge: | Existing research on stereotypical biases ignores literature on them and results in resource wastage. |
| Approach: | They argue that stereotypes are social constructs shaping human perception and behavior that can produce harmful outcomes under specific conditions. |
| Outcome: | The proposed models can inherit and amplify stereotypes under certain conditions. |
Cognitive Effects and Biases in Large Language Models (2026.eacl-tutorials)
Copied to clipboard
| Challenge: | This tutorial bridges psychology and NLP to clarify cognitive effects and biases in large language models. |
| Approach: | This tutorial bridges psychology and NLP to clarify cognitive effects and biases in large language models. |
| Outcome: | This tutorial bridges psychology and NLP to clarify cognitive effects and biases in large language models. |
On Measures of Biases and Harms in NLP (2022.findings-aacl)
Copied to clipboard
Sunipa Dev, Emily Sheng, Jieyu Zhao, Aubrie Amstutz, Jiao Sun, Yu Hou, Mattie Sanseverino, Jiin Kim, Akihiro Nishi, Nanyun Peng, Kai-Wei Chang
| Challenge: | Recent studies show that natural language processing (NLP) technologies propagate societal biases about demographic groups associated with attributes such as gender, race, and nationality. |
| Approach: | They propose a framework for harms and questions to help practitioners understand biases . they propose measurable measures to detect and mitigate biased groups . |
| Outcome: | The proposed framework provides a framework for harms and questions for practitioners to answer to guide the development of bias measures. |
From Prejudice to Parity: A New Approach to Debiasing Large Language Model Word Embeddings (2025.coling-main)
Copied to clipboard
| Challenge: | Existing work in this field has looked most commonly into gender bias, racial bias, and religious bias. |
| Approach: | They propose an algorithm that uses a neural network to perform ‘soft debiasing’ and build on the seminal work of (CITATION) and (CitATION). |
| Outcome: | The proposed algorithm outperforms current methods on gender, race, and religion metrics on a wide range of metrics. |
Social Bias in Multilingual Language Models: A Survey (2025.emnlp-main)
Copied to clipboard
| Challenge: | Pretrained multilingual models exhibit the same social bias as models processing English texts. |
| Approach: | They examine the literature on bias evaluation and mitigation approaches in multilingual and non-English contexts and identify gaps in the field. |
| Outcome: | The proposed models perform well on multilingual language understanding benchmarks and are consistent with the current literature. |
AccessEval: Benchmarking Disability Bias in Large Language Models (2025.emnlp-main)
Copied to clipboard
| Challenge: | Large Language Models exhibit disparities in how they handle real life queries. |
| Approach: | They propose a large-scale benchmark to evaluate large language models across six real-world domains and nine disability types. |
| Outcome: | The proposed model outputs show higher factual error, more negative tone, and increased stereotyping with social perception compared to neutral queries. |