Papers by Katerina Korre

5 papers
Untangling Hate Speech Definitions: A Semantic Componential Analysis Across Cultures and Domains (2025.findings-naacl)

Copied to clipboard

Challenge: a new framework for analyzing hate speech definitions is proposed to address cultural differences in interpretations . a dataset of 493 definitions from more than 100 cultures is used to analyze hate speech .
Approach: They propose a framework for a cross-cultural and cross-domain analysis of hate speech definitions . they use open-source LLMs to analyze the impact of different definitions on hate speech detection .
Outcome: The proposed framework enables cross-cultural and cross-domain analysis of hate speech definitions . it reveals that many domains borrow definitions from one another without taking into account target culture .
A Corpus for Sentence-Level Subjectivity Detection on English News Articles (2024.lrec-main)

Copied to clipboard

Challenge: Existing approaches to spotting subjectivity require language-specific tools.
Approach: They develop annotation guidelines for sentence-level subjectivity detection that are not limited to language-specific cues.
Outcome: The proposed framework enables subjectivity detection in English and across other languages without relying on language-specific tools, such as lexicons or machine translation.
The Challenges of Creating a Parallel Multilingual Hate Speech Corpus: An Exploration (2024.lrec-main)

Copied to clipboard

Challenge: Hate speech is one of the most demanding topics in Natural Language Processing, as its multifaceted nature is accompanied by a handful of challenges, such as multilinguality and cross-linguality.
Approach: They propose a pipeline that could be used to create a parallel multilingual hate speech dataset using machine translation.
Outcome: The proposed pipeline will be able to create a parallel multilingual hate speech dataset using machine translation.
Enriching Grammatical Error Correction Resources for Modern Greek (2022.lrec-1)

Copied to clipboard

Challenge: Davidson and Kilgarriff, 2011) have focused on the English language, but there are limited efforts to expand GEC in other languages.
Approach: They develop and test a multilingual text-to-text transformer for Greek . they provide a model that can be fully-fledged for Greek with annotation corrections .
Outcome: The proposed model achieves 52.63% F0.5 on part of the Greek Native Corpus, 16% below the winning system on English GEC.
Evaluation and Facilitation of Online Discussions in the LLM Era: A Survey (2025.emnlp-main)

Copied to clipboard

Challenge: Recent advances in LLMs enable artificial facilitation agents to not only moderate content, but also actively improve the quality of interactions.
Approach: They propose a taxonomy on discussion quality evaluation and a new taxonomies for intervention and facilitation strategies.
Outcome: The proposed methods synthesize ideas from Natural Language Processing (NLP) and Social Sciences to provide a taxonomy on discussion quality evaluation, and a roadmap of good practices and future research directions.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations