Quantifying the Semantic Core of Gender Systems (D19-1)

Copied to clipboard

Challenge: a large number of languages employ grammatical gender on the lexeme, but is it truly arbitrary? a recent study shows that the relationship between grammamatical gender and lexical semantics is opaque.
Approach: They propose a method to correlating inanimate nouns' gender with lexical semantics . they find that the gender systems of 18 languages exhibit a significant correlation with a definition .
Outcome: a new study shows that the gender assignments of 18 languages are arbitrary . the authors show that the correlation between gender and semantics is significant .

Similar Papers

Measuring the Similarity of Grammatical Gender Systems by Comparing Partitions (2020.emnlp-main)

Copied to clipboard

Challenge: A grammatical gender system divides a lexicon into a small number of fixed categories with fixed usage across speakers.
Approach: They propose to define gender systems extensionally to reduce comparisons to cluster evaluation by comparing pairwise overlaps between gender systems.
Outcome: The proposed measures are based on a phylogenetic tree over extant Indo-European languages.
On the Relationships Between the Grammatical Genders of Inanimate Nouns and Their Co-Occurring Adjectives and Verbs (2021.tacl-1)

Copied to clipboard

Challenge: In many languages, nouns possess grammatical genders.
Approach: They use large-scale corpora and tools from NLP and information theory to test whether there is a relationship between grammatical genders of inanimate nouns and adjectives used to describe them.
Outcome: The results show that there is a statistically significant relationship between the grammatical genders of inanimate nouns and adjectives used to describe them in all six languages.
Unsupervised Discovery of Gendered Language through Latent-Variable Modeling (P19-1)

Copied to clipboard

Challenge: a recent study has focused on the ways in which language is gendered . positive adjectives used to describe women are more often related to their bodies .
Approach: They propose a model that models adjective choice and its sentiment given the natural gender of a head noun.
Outcome: The proposed model shows that positive adjectives used to describe women are more often related to their bodies than positive adjective words used to explain men.
Examining Gender Bias in Languages with Grammatical Gender (D19-1)

Copied to clipboard

Challenge: Existing studies on gender bias in word embeddings focus on English . however, these studies cannot be extended to languages with morphological agreement on gender .
Approach: They propose new metrics to evaluate gender bias in word embeddings of English and Spanish . they extend existing approaches to mitigate gender bias while preserving original embeddables .
Outcome: The proposed methods reduce gender bias while preserving the original embeddings.
Pick a Fight or Bite your Tongue: Investigation of Gender Differences in Idiomatic Language Usage (2020.coling-main)

Copied to clipboard

Challenge: Existing studies on gender-linked language have established foundations regarding cross-gender differences in lexical, emotional, and topical preferences, along with their sociological underpinnings.
Approach: They compile a corpus of spontaneous linguistic productions annotated with speakers’ gender and perform an empirical study of gender differences in the usage of figurative language between male and female authors.
Outcome: The results show that gender-specific idiomatic choices reflect gender-based lexical and semantic preferences in general language, men's and women's idioms express higher emotion than their literal language, and contextual analysis of idiomatic expressions reveals considerable differences, reflecting subtle divergences in usage environments, shaped by cross-gender communication styles and semantic biases.
Toward Gender-Inclusive Coreference Resolution (2020.acl-main)

Copied to clipboard

Challenge: a recent study shows that coreference resolution systems can be harmful to binary and non-binary trans and cis stakeholders.
Approach: They propose to use gender-based crowd annotations to investigate coreference resolution biases . they use a dataset to examine the complexity of gender in crowd annotation systems .
Outcome: a new study shows that without acknowledging and building systems that recognize gender, we build systems that lead to many potential harms.
What social attitudes about gender does BERT encode? Leveraging insights from psycholinguistics (2023.acl-long)

Copied to clipboard

Challenge: Much research has focused on evaluating whether large language models encode stereotypical/harmful associations.
Approach: They propose to use two datasets from human experiments to examine how word preferences in a large language model reflect social attitudes about gender.
Outcome: The language model BERT takes into account factors that shape human lexical choice of such language, but may not weigh those factors in the same way people do.
GenderQuant: Quantifying Mention-Level Genderedness (N19-1)

Copied to clipboard

Challenge: Existing approaches to detect gendered language require considerable annotation efforts for each language, domain, and author, and often require handcrafted lexicons and features.
Approach: They use existing NLP pipelines to automatically annotate gender of mentions in the text and train a supervised classifier to predict the gender of any mention from its context and evaluate it on unseen text.
Outcome: The proposed method can detect gendered language on movie summaries, movie reviews, news articles, and fiction novels.
Mandarin classifier systems optimize to accommodate communicative pressures (2023.findings-emnlp)

Copied to clipboard

Challenge: Existing studies suggest that gendered languages are inherently optimized to accommodate communicative pressures on language learning and processing.
Approach: They propose to use grammatical or probabilistic modifiers to smooth the entropy of nouns in context to find the same frequency, similarity, and co-occurrence interactions that structure gender systems.
Outcome: The proposed noun classification device is sensitive to frequency, similarity, and co-occurrence interactions that structure gender systems.
Under the Morphosyntactic Lens: A Multifaceted Evaluation of Gender Bias in Speech Translation (2022.acl-long)

Copied to clipboard

Challenge: grammatical gender languages are characterized by morphosyntactic chains of gender agreement marked on a variety of lexical items and parts-of-speech (POS).
Approach: They propose to enrich the natural, gender-sensitive MuST-SHE corpus with two new linguistic annotation layers to explore gender bias.
Outcome: The proposed models shed light on gender bias and its detection at several levels of granularity.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations