Analyzing Norm Violations in Live-Stream Chat (2023.emnlp-main)

Copied to clipboard

Challenge: Existing methods for detecting toxic language and norm violations are limited to live-streaming platforms . existing methods are less effective when applied to live streaming platforms based on a limited time frame .
Approach: They propose to use contextual information to automatically moderate toxic content on live streaming platforms.
Outcome: The proposed model improves on live-streaming platforms by 35%.

Similar Papers

Detecting Community Sensitive Norm Violations in Online Conversations (2021.findings-emnlp)

Copied to clipboard

Challenge: Existing efforts to identify unacceptable behavior have focused on toxicity as the sole form of community norm violation.
Approach: They propose a dataset that focuses on a more complete spectrum of community norms and their violations in local conversational and global contexts.
Outcome: The proposed model improves the detection of community norm violations in local conversational and global contexts.
Offensive Language Detection on Video Live Streaming Chat (2020.coling-main)

Copied to clipboard

Challenge: a prototype of a live chat room that detects offensive expressions in live streaming chats is presented . offensive expression detection on social media platforms can provide more protection for users .
Approach: They propose a live chat room that detects offensive expressions in live streaming chats in real time . they used a dataset from Twitch to analyze offensive expression patterns .
Outcome: The proposed chat room detects offensive expressions in live streaming chats in real time.
A Stacking-based Efficient Method for Toxic Language Detection on Live Streaming Chat (2022.emnlp-industry)

Copied to clipboard

Challenge: Existing methods for toxic language detection are based on deep learning, but they are not scalable considering inference speed and computational resources.
Approach: They propose a method for toxic language detection that is aware of real-world scenarios by partial stacking partial stacks that feeds initial results with low confidence to meta-classifier.
Outcome: The proposed method achieves faster inference speed than BERT-based models with comparable performance.
LLM-Human Pipeline for Cultural Grounding of Conversations (2025.naacl-long)

Copied to clipboard

Challenge: addressing parents by name is commonplace in the West, but it is rare in most Asian cultures.
Approach: They propose a Cultural Context Schema for conversations that incorporates conversational information and cultural information such as social norms, violations, etc.
Outcome: The proposed model significantly improves the empirical performance of a Chinese conversational norm and violation description using an interactive human-in-loop framework.
Silencing Empowerment, Allowing Bigotry: Auditing the Moderation of Hate Speech on Twitch (2025.acl-long)

Copied to clipboard

Challenge: To meet the demands of content moderation, online platforms have resorted to automated systems.
Approach: They conduct an audit of Twitch’s automated moderation tool (AutoMod) to investigate its effectiveness in flagging hateful content.
Outcome: The automated moderation tool (AutoMod) is used to filter hateful content on Twitch and send 107,000 comments from 4 datasets.
RENOVI: A Benchmark Towards Remediating Norm Violations in Socio-Cultural Conversations (2024.findings-naacl)

Copied to clipboard

Challenge: Norm violations occur when individuals fail to conform to culturally accepted behaviors, which may lead to potential conflicts.
Approach: They propose to use a large corpus of 9,258 multi-turn dialogues annotated with social norms to equip AI systems with a remediation ability.
Outcome: The proposed system can understand and remediate norm violations step by step.
NormDial: A Comparable Bilingual Synthetic Dialog Dataset for Modeling Social Norm Adherence and Violation (2023.emnlp-main)

Copied to clipboard

Challenge: Social norms fundamentally shape interpersonal communication.
Approach: They propose a human-in-the-loop pipeline to synthesize a bilingual dyadic dialogue dataset with turn-by-turn annotations of social norms for Chinese and American cultures.
Outcome: The proposed dataset is high-quality through human evaluation and compares with existing models.
NORMSAGE: Multi-Lingual Multi-Cultural Norm Discovery from Conversations On-the-Fly (2023.emnlp-main)

Copied to clipboard

Challenge: Existing methods to understand acceptable behavior have focused on a single culture and manually built datasets from non-conversational settings.
Approach: They propose a framework to automatically extract culture-specific norms from multi-lingual conversations.
Outcome: The proposed framework extracts culture-specific norms from multi-lingual conversations.
PluRule: A Benchmark for Moderating Pluralistic Communities on Social Media (2026.acl-long)

Copied to clipboard

Challenge: Social media are shifting towards community-governed platforms where groups define their own norms.
Approach: They propose a multimodal, multilingual benchmark for detecting 13,371 rule violations across 1,989 Reddit communities . they show that bigger models and increased context provide marginal gains, and universal rules like civility and self-promotion are easier to detect.
Outcome: The proposed model can detect 13,371 rule violations across 1,989 Reddit communities across 2,885 rules in 9 languages.
ToxicChat: Unveiling Hidden Challenges of Toxicity Detection in Real-World User-AI Conversation (2023.findings-emnlp)

Copied to clipboard

Challenge: toxicity detection has been largely based on social media content, leaving the unique challenges inherent to real-world user-AI interactions insufficiently explored.
Approach: They propose a benchmark to detect toxicity in real-world user-AI conversations . they compare existing models with social media content to find toxicity .
Outcome: The proposed benchmark reveals that existing models fail to recognize toxicity in real-world user-AI conversations.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations