Papers by Sravani Nanduri
BiasX: “Thinking Slow” in Toxic Content Moderation with Explanations of Implied Social Biases (2023.emnlp-main)
Copied to clipboard
| Challenge: | Toxicity annotators and content moderators often default to mental shortcuts when making decisions, leading to subtle toxicity being missed and seemingly harmless content being over-detected. |
| Approach: | They propose a framework that provides AI-generated explanations of statements’ implied social biases to enhance content moderation setups. |
| Outcome: | The proposed framework significantly improves content moderation setups by enabling users to think more thoroughly about their decisions. |