Papers by Katerina Korre
Untangling Hate Speech Definitions: A Semantic Componential Analysis Across Cultures and Domains (2025.findings-naacl)
Copied to clipboard
| Challenge: | a new framework for analyzing hate speech definitions is proposed to address cultural differences in interpretations . a dataset of 493 definitions from more than 100 cultures is used to analyze hate speech . |
| Approach: | They propose a framework for a cross-cultural and cross-domain analysis of hate speech definitions . they use open-source LLMs to analyze the impact of different definitions on hate speech detection . |
| Outcome: | The proposed framework enables cross-cultural and cross-domain analysis of hate speech definitions . it reveals that many domains borrow definitions from one another without taking into account target culture . |
A Corpus for Sentence-Level Subjectivity Detection on English News Articles (2024.lrec-main)
Copied to clipboard
Francesco Antici, Federico Ruggeri, Andrea Galassi, Katerina Korre, Arianna Muti, Alessandra Bardi, Alice Fedotova, Alberto Barrón-Cedeño
| Challenge: | Existing approaches to spotting subjectivity require language-specific tools. |
| Approach: | They develop annotation guidelines for sentence-level subjectivity detection that are not limited to language-specific cues. |
| Outcome: | The proposed framework enables subjectivity detection in English and across other languages without relying on language-specific tools, such as lexicons or machine translation. |
The Challenges of Creating a Parallel Multilingual Hate Speech Corpus: An Exploration (2024.lrec-main)
Copied to clipboard
| Challenge: | Hate speech is one of the most demanding topics in Natural Language Processing, as its multifaceted nature is accompanied by a handful of challenges, such as multilinguality and cross-linguality. |
| Approach: | They propose a pipeline that could be used to create a parallel multilingual hate speech dataset using machine translation. |
| Outcome: | The proposed pipeline will be able to create a parallel multilingual hate speech dataset using machine translation. |
Enriching Grammatical Error Correction Resources for Modern Greek (2022.lrec-1)
Copied to clipboard
| Challenge: | Davidson and Kilgarriff, 2011) have focused on the English language, but there are limited efforts to expand GEC in other languages. |
| Approach: | They develop and test a multilingual text-to-text transformer for Greek . they provide a model that can be fully-fledged for Greek with annotation corrections . |
| Outcome: | The proposed model achieves 52.63% F0.5 on part of the Greek Native Corpus, 16% below the winning system on English GEC. |
Evaluation and Facilitation of Online Discussions in the LLM Era: A Survey (2025.emnlp-main)
Copied to clipboard
Katerina Korre, Dimitris Tsirmpas, Nikos Gkoumas, Emma Cabalé, Danai Myrtzani, Theodoros Evgeniou, Ion Androutsopoulos, John Pavlopoulos
| Challenge: | Recent advances in LLMs enable artificial facilitation agents to not only moderate content, but also actively improve the quality of interactions. |
| Approach: | They propose a taxonomy on discussion quality evaluation and a new taxonomies for intervention and facilitation strategies. |
| Outcome: | The proposed methods synthesize ideas from Natural Language Processing (NLP) and Social Sciences to provide a taxonomy on discussion quality evaluation, and a roadmap of good practices and future research directions. |