Challenge: Recent studies have found evidence of gender bias in machine translation and coreference resolution models using mostly synthetic diagnostic datasets.
Approach: They propose a semi-automatic method to vastly extend synthetic, small diagnostic datasets to include grammatical patterns indicating stereotypical and non-stereotypical gender-role assignments.
Outcome: The proposed method extends the existing dataset to 108K diverse English sentences.

Similar Papers

Evaluating Gender Bias in Machine Translation (P19-1)

Copied to clipboard

Challenge: Using morphological analysis, we find that MT models exhibit gender-biased translation errors when training data encode stereotypes not relevant for the task.
Approach: They propose an automatic gender bias evaluation method for eight target languages with grammatical gender based on morphological analysis.
Outcome: The proposed method is based on two recent coreference resolution datasets composed of English sentences cast participants into non-stereotypical gender roles.
Gender Bias in Coreference Resolution (N18-2)

Copied to clipboard

Challenge: a study of coreference resolution systems that resolve gender differences in pairs is aimed at examining implicit gender biases.
Approach: They propose a Winograd schema-style set of minimal pair sentences that differ only by gender . they evaluate and confirm systematic gender bias in three publicly-available coreference resolution systems .
Outcome: The proposed system resolves a male and neutral pronoun as coreferent with "The surgeon" but does not resolve the female pronounce.
Toward Gender-Inclusive Coreference Resolution (2020.acl-main)

Copied to clipboard

Challenge: a recent study shows that coreference resolution systems can be harmful to binary and non-binary trans and cis stakeholders.
Approach: They propose to use gender-based crowd annotations to investigate coreference resolution biases . they use a dataset to examine the complexity of gender in crowd annotation systems .
Outcome: a new study shows that without acknowledging and building systems that recognize gender, we build systems that lead to many potential harms.
Under the Morphosyntactic Lens: A Multifaceted Evaluation of Gender Bias in Speech Translation (2022.acl-long)

Copied to clipboard

Challenge: grammatical gender languages are characterized by morphosyntactic chains of gender agreement marked on a variety of lexical items and parts-of-speech (POS).
Approach: They propose to enrich the natural, gender-sensitive MuST-SHE corpus with two new linguistic annotation layers to explore gender bias.
Outcome: The proposed models shed light on gender bias and its detection at several levels of granularity.
Does Context Help Mitigate Gender Bias in Neural Machine Translation? (2024.findings-emnlp)

Copied to clipboard

Challenge: Neural machine translation models perpetuate gender bias in their training data distribution.
Approach: They examine the gender bias in Neural Machine Translation by using context-aware models to enhance translation accuracy for feminine terms and translation with non-informative context in Basque to Spanish.
Outcome: The proposed models can maintain or even amplify gender bias in translations of stereotypical professions in English and with non-informative context in Basque to Spanish.
Gender Bias in Coreference Resolution: Evaluation and Debiasing Methods (N18-2)

Copied to clipboard

Challenge: Existing methods for co-reference resolution focus on gender bias.
Approach: They propose a new benchmark for co-reference resolution focused on gender bias, WinoBias.
Outcome: The proposed system removes the bias without significantly affecting performance on existing datasets.
Identifying and Reducing Gender Bias in Word-Level Language Models (N19-3)

Copied to clipboard

Challenge: Existing discriminatory biases in training data can be amplified by models . text corpora exhibit socially problematic biase .
Approach: They propose a metric to measure gender bias and a regularization loss term to minimize embeddings onto an embeddable subspace that encodes gender.
Outcome: The proposed method reduces gender bias up to an optimal weight assigned to the loss term, and the model becomes unstable as the perplexity increases.
How sensitive are translation systems to extra contexts? Mitigating gender bias in Neural Machine Translation models through relevant contexts. (2022.findings-emnlp)

Copied to clipboard

Challenge: Neural Machine Translation systems are prone to gender biases in their learned representations.
Approach: They propose to use contextual sentences to correct gender bias in Neural Machine Translation models.
Outcome: The proposed method can be used to build better, bias-free translation systems.
Multi-Dimensional Gender Bias Classification (2020.emnlp-main)

Copied to clipboard

Challenge: a novel framework decomposes gender bias in text along several pragmatic and semantic dimensions . language is a primary means by which people communicate, express identities and categorize themselves . unwanted gender biases can affect downstream applications, leading to poor user experiences .
Approach: They propose a framework that decomposes gender bias in text along several dimensions . they annotate eight large scale datasets with gender information and collect a benchmark .
Outcome: The proposed framework decomposes gender bias in text along several pragmatic and semantic dimensions.
Mitigating Gender Bias in Natural Language Processing: Literature Review (P19-1)

Copied to clipboard

Challenge: NLP models propagate and may even amplify gender bias found in text corpora . methods to mitigate gender bias in NLP are relatively nascent .
Approach: They propose to analyze gender bias based on four forms of representation bias and discuss the advantages and drawbacks of existing gender debiasing methods.
Outcome: The proposed methods are based on four forms of representation bias and have advantages and drawbacks.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations