Challenge: Existing studies on mental manipulation focus on context-free content and face challenges in identifying implicit toxicity.
Approach: They propose a dataset that analyzes mental manipulation and its components . they propose to use 4,000 fictional dialogues to identify the techniques utilized for manipulation .
Outcome: The proposed dataset enables a comprehensive analysis of mental manipulation . it shows that leading-edge models inadequately identify and categorize manipulative content .

Similar Papers

Detecting Conversational Mental Manipulation with Intent-Aware Prompting (2025.coling-main)

Copied to clipboard

Challenge: Existing approaches to detect mental manipulations are limited due to complexity of detecting subtle, covert tactics in conversations.
Approach: They propose an approach to detect mental manipulations using large language models using intent-aware prompting by capturing the intents of participants.
Outcome: The proposed approach significantly reduces false negatives, helping detect more instances of mental manipulation with minimal misjudgment of positive cases.
SELF-PERCEPT: Introspection Improves Large Language Models’ Detection of Multi-Person Mental Manipulation in Conversations (2025.acl-short)

Copied to clipboard

Challenge: Mental manipulation is subtle yet pervasive form of abuse in interpersonal communication, making its detection critical for safeguarding potential victims.
Approach: They propose a dataset of 220 multi-turn, multi-person dialogues balanced between manipulative and non-manipulative interactions drawn from reality shows that mimic real-life scenarios.
Outcome: The proposed framework shows that it can detect multi-person, multi-turn mental manipulation in multi-people conversations.
Preparing Data from Psychotherapy for Natural Language Processing (L18-1)

Copied to clipboard

Challenge: mental health care is a demanding occupation, resulting in a severe gap in patient-centered care . a recent study shows that natural language processing can extract certain aspects of human-human communication.
Approach: They propose to use data from psychotherapy sessions to help improve quality of care . they use feedback and cooperation annotations to assess quality of therapy sessions .
Outcome: The proposed method aims to analyse psychotherapy data and assess its quality . it aims at identifying what qualifies for good feedback or cooperation in therapy sessions .
A Survey of Cognitive Distortion Detection and Classification in NLP (2025.findings-emnlp)

Copied to clipboard

Challenge: despite momentum in natural language processing, the field remains fragmented . inconsistencies in CD taxonomies, task formulations and evaluation practices limit comparability .
Approach: This review provides a comprehensive review of 38 studies spanning two decades . they map how CDs have been implemented in computational research and evaluate the methods applied.
Outcome: The paper presents the first comprehensive review of 38 studies spanning two decades . it summarises common task setups and highlights persistent challenges to support more coherent research.
ManiTweet: A New Benchmark for Identifying Manipulation of News on Social Media (2025.coling-main)

Copied to clipboard

Challenge: Existing studies have focused on the identification of social media posts that contain misrepresentations of information within associated news articles.
Approach: They propose a data collection schema and curated a dataset called ManiTweet, consisting of 3.6K pairs of tweets and corresponding articles.
Outcome: The proposed model outperforms large language models on the ManiTweet dataset and reveals intriguing connections between manipulation and the domain and factuality of news articles.
ConvAbuse: Data, Analysis, and Benchmarks for Nuanced Abuse Detection in Conversational AI (2021.emnlp-main)

Copied to clipboard

Challenge: Existing studies on abusive language towards conversational AI systems are not conclusive as they are not performed with live systems nor with real users due to the lack of reliable abuse detection tools.
Approach: They propose to use a convAI dataset to account for the complexity of the task and to bench-mark existing models against this data.
Outcome: The proposed model shows that abuse distribution is different compared to other datasets, with sexual tinted aggression towards the virtual persona of the systems.
Introducing CAD: the Contextual Abuse Dataset (2021.naacl-main)

Copied to clipboard

Challenge: Detecting and classifying online abuse is a complex and nuanced task, despite many advances in the power and availability of computational tools.
Approach: They propose to annotate a reddit conversation thread with six distinct primary and secondary categories and an expert-driven group-adjudication process for high quality annotations.
Outcome: The proposed dataset contains six distinct primary and secondary categories and uses an expert-driven group-adjudication process for high quality annotations.
MentalHelp: A Multi-Task Dataset for Mental Health in Social Media (2024.lrec-main)

Copied to clipboard

Challenge: Annotating social media data for mental health disorders is expensive and time-consuming, limiting their size and scope.
Approach: They present a large-scale semi-supervised mental disorder detection dataset containing 14 million instances from Reddit and an ensemble of three separate models.
Outcome: The proposed dataset contains 14 million instances of mental disorders . it was collected from reddit and labeled in a semi-supervised way .
Joint Modelling of Emotion and Abusive Language Detection (2020.acl-main)

Copied to clipboard

Challenge: Existing methods for abuse detection focus on linguistic properties of comments and online communities of users, disregarding the emotional state of the users and how this might affect their language.
Approach: They propose to combine emotion and abusive language detection to create a multi-task learning framework that allows one task to inform the other.
Outcome: The proposed model improves on the previous models, incorporating affective features into the learning framework.
A Just and Comprehensive Strategy for Using NLP to Address Online Abuse (P19-1)

Copied to clipboard

Challenge: Current methods to detect online abuse focus on a narrow definition of abuse to detriment of victims seeking validation and solutions.
Approach: They argue that the NLP community needs to make three substantive changes to tackle both more subtle and more serious forms of abuse.
Outcome: The proposed approach would address the problem of abuse in a more inclusive and productive way.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations