Uncover Sexual Harassment Patterns from Personal Stories by Joint Key Element Extraction and Categorization (D19-1)
Copied to clipboard
| Challenge: | Sexual harassment is a pervasive, worldwide problem with a long history . statistics show that girls and women are put at high risk of experiencing harassment. |
| Approach: | They manually annotated sexual harassment stories with labels in dimensions of location, time, and harassers’ characteristics and applied natural language processing techniques to extract key elements at the same time. |
| Outcome: | The proposed algorithms will help people who have been harassed, authorities, researchers and other related parties in various ways, such as automatically filling reports, and enabling faster action to be taken. |
Similar Papers
SafeCity: Understanding Diverse Forms of Sexual Harassment Personal Stories (D18-1)
Copied to clipboard
| Challenge: | With the recent rise of #MeToo, an increasing number of personal stories about sexual harassment and sexual abuse have been shared online. |
| Approach: | They propose to use CNN-RNN model to automatically categorize and analyze sexual harassment data from SafeCity forums. |
| Outcome: | The proposed model achieves an accuracy of 86.5% for groping, ogling, and commenting, and 82.5% in multi-label models. |
#YouToo? Detection of Personal Recollections of Sexual Harassment on Social Media (P19-1)
Copied to clipboard
| Challenge: | a recent study has found that the disclosure of sexual abuse has positive psychological im- pacts. |
| Approach: | They propose to aggregate personal experiences of sexual harassment from Twitter posts to facilitate a better understanding of social media constructs and bring about social change. |
| Outcome: | The proposed model is compared with state-of-the-art models and is based on a three part Twitter-Specific Social Media Language Model. |
Speak up, Fight Back! Detection of Social Media Disclosures of Sexual Harassment (N19-3)
Copied to clipboard
| Challenge: | #MeToo movement provides platform to narrate personal experiences of sexual harassment. |
| Approach: | They propose a three-part ULMFiT architecture to tackle text subtleties in a classification task . they propose to annotate a manually annotated real-world dataset to test their approach . |
| Outcome: | The proposed model outperforms existing models that rely on handcrafted stylistic features and is more accurate than generic models. |
Reports of personal experiences and stories in argumentation: datasets and analysis (2022.acl-long)
Copied to clipboard
| Challenge: | Personal experiences and stories are important in argumentation, but they are not considered in the social sciences. |
| Approach: | They propose to use annotated documents to scale-up the analysis using existing annotations. |
| Outcome: | The proposed classifiers can identify documents containing personal experiences and reports . they can scale up to three domains and show that they perform well across domains. |
Introducing CAD: the Contextual Abuse Dataset (2021.naacl-main)
Copied to clipboard
| Challenge: | Detecting and classifying online abuse is a complex and nuanced task, despite many advances in the power and availability of computational tools. |
| Approach: | They propose to annotate a reddit conversation thread with six distinct primary and secondary categories and an expert-driven group-adjudication process for high quality annotations. |
| Outcome: | The proposed dataset contains six distinct primary and secondary categories and uses an expert-driven group-adjudication process for high quality annotations. |
Multitask Learning for Emotionally Analyzing Sexual Abuse Disclosures (2021.naacl-main)
Copied to clipboard
| Challenge: | Prior work on identifying narratives related to sexual abuse disclosures did not consider this as an independent task. |
| Approach: | They propose to identify narratives related to sexual abuse disclosures as a joint modeling task that leverages their emotional attributes through multitask learning. |
| Outcome: | The proposed model leverages emotional attributes of textual conversations to identify narratives related to sexual abuse disclosures in homogeneous and heterogeneously settings. |
Annotating Online Misogyny (2021.acl-long)
Copied to clipboard
| Challenge: | Online misogyny is a category of online abusive language with serious and harmful social consequences. |
| Approach: | They propose an iterative annotation process and a taxonomy of labels for annotating misogyny in natural written language and cite a high-quality dataset of annotated posts from social media posts. |
| Outcome: | The proposed method aims to identify misogynistic language in natural written language and annotate it in social media posts using a high-quality dataset. |
An Expert Annotated Dataset for the Detection of Online Misogyny (2021.eacl-main)
Copied to clipboard
| Challenge: | Existing studies have found that misogynistic content is pervasive on some Reddit communities, but a training dataset for misogorical classification has not been created with the data. |
| Approach: | They propose a hierarchical taxonomy and an expert labelled dataset to enable automatic classification of online misogynistic content. |
| Outcome: | The proposed taxonomy and an expert labelled dataset are made freely available for future research. |
MentalManip: A Dataset For Fine-grained Analysis of Mental Manipulation in Conversations (2024.acl-long)
Copied to clipboard
| Challenge: | Existing studies on mental manipulation focus on context-free content and face challenges in identifying implicit toxicity. |
| Approach: | They propose a dataset that analyzes mental manipulation and its components . they propose to use 4,000 fictional dialogues to identify the techniques utilized for manipulation . |
| Outcome: | The proposed dataset enables a comprehensive analysis of mental manipulation . it shows that leading-edge models inadequately identify and categorize manipulative content . |
Author Profiling for Abuse Detection (C18-1)
Copied to clipboard
| Challenge: | Existing methods for detecting abusive content rely on textual cues and lexical cue information. |
| Approach: | They propose a method that incorporates community-based profiling features of Twitter users to detect abusive content by using a dataset of 16k tweets. |
| Outcome: | The proposed approach outperforms the current state-of-the-art in abuse detection on a dataset of 16k tweets. |