Papers by Patricia Rossini
“It’s Not Just Hate”: A Multi-Dimensional Perspective on Detecting Harmful Speech Online (2022.emnlp-main)
Copied to clipboard
| Challenge: | Detecting offensive content is becoming a critical task in natural language processing . but most datasets use a single binary label for hate or incivility, even though each concept is multi-faceted . a more fine-grained multi-label approach addresses conceptual and performance issues . |
| Approach: | They propose to use a dataset to annotate offensive online speech with six labels . they propose to apply a more fine-grained approach to predicting incivility and hateful content . |
| Outcome: | The proposed approach outperforms or matches benchmark datasets on the annotated tweets. |
Introducing CAD: the Contextual Abuse Dataset (2021.naacl-main)
Copied to clipboard
| Challenge: | Detecting and classifying online abuse is a complex and nuanced task, despite many advances in the power and availability of computational tools. |
| Approach: | They propose to annotate a reddit conversation thread with six distinct primary and secondary categories and an expert-driven group-adjudication process for high quality annotations. |
| Outcome: | The proposed dataset contains six distinct primary and secondary categories and uses an expert-driven group-adjudication process for high quality annotations. |