Papers by Jiangrui Zheng
HateModerate: Testing Hate Speech Detectors against Content Moderation Policies (2024.findings-naacl)
Copied to clipboard
| Challenge: | Existing studies on hate speech detection have failed to answer this question. |
| Approach: | They propose a dataset for testing the behaviors of automated content moderators against content policies. |
| Outcome: | The proposed dataset includes hateful and non-hateful examples matching the 41 community standards guideline policies of Facebook. |