Papers by Hezhao Zhang
Beyond Hate Speech: NLP’s Challenges and Opportunities in Uncovering Dehumanizing Language (2025.emnlp-main)
Copied to clipboard
| Challenge: | Existing hate speech datasets rarely contain enough instances of dehumanizing content, and current models struggle to distinguish such language from more benign forms of hate or offense. |
| Approach: | They evaluate four state-of-the-art large language models for dehumanization detection. |
| Outcome: | The proposed models perform only moderately under an optimized configuration, while others over-predict dehumanization for some identities, while under-identifying it for others. |