Papers by Daniel Licht

2 papers
Multilingual Holistic Bias: Extending Descriptors and Patterns to Unveil Demographic Biases in Languages at Scale (2023.emnlp-main)

Copied to clipboard

Challenge: Multilingual HolisticBias dataset includes 20,459 sentences in 50 languages . dataset is intended to uncover demographic imbalances and quantify mitigations .
Approach: They propose a multilingual extension of the HolisticBias dataset . they use 118 demographic descriptors and three patterns to build multilingual sentences .
Outcome: The proposed model improves translation quality when the source input only differs in gender . it also improves when the masculine human reference is used in the model .
Toxicity in Multilingual Machine Translation at Scale (2023.findings-emnlp)

Copied to clipboard

Challenge: In this paper, we evaluate and analyze added toxicity when translating a large dataset from English into 164 languages.
Approach: They evaluate added toxicity when translating a large dataset from English into 164 languages.
Outcome: The results show that added toxicity is more prevalent in low-resource languages than in high-resolution translations.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations