PolyNarrative: A Multilingual, Multilabel, Multi-domain Dataset for Narrative Extraction from News Articles (2025.acl-long)
Copied to clipboard
Nikolaos Nikolaidis, Nicolas Stefanovitch, Purificação Silvano, Dimitar Iliyanov Dimitrov, Roman Yangarber, Nuno Guimarães, Elisa Sartori, Ion Androutsopoulos, Preslav Nakov, Giovanni Da San Martino, Jakub Piskorski
| Challenge: | a new dataset of news articles annotated for narratives provides a framework for narrative detection . recurring narratives can propagate with very high velocity across audiences, languages and countries . |
| Approach: | They propose a multilingual dataset annotated for narratives using two-level taxonomies . they define narrative as a recurring, repetitive, overt or implicit claim that promotes a specific interpretation or viewpoint on an ongoing topic . |
| Outcome: | The proposed dataset will foster research in narrative detection and enable new research directions . the authors identify multiple narratives in the same article, and the results are published online . |
Similar Papers
Entity Framing and Role Portrayal in the News (2025.findings-acl)
Copied to clipboard
Tarek Mahmoud, Zhuohan Xie, Dimitar Iliyanov Dimitrov, Nikolaos Nikolaidis, Purificação Silvano, Roman Yangarber, Shivam Sharma, Elisa Sartori, Nicolas Stefanovitch, Giovanni Da San Martino, Jakub Piskorski, Preslav Nakov
| Challenge: | a dataset of news articles containing 22 fine-grained characters is annotated for entity framing and role portrayal . the dataset includes 1,378 recent news articles in five languages focusing on the Ukraine-Russia War and climate change . |
| Approach: | They propose a multilingual and hierarchical corpus annotated for entity framing and role portrayal in news articles. |
| Outcome: | The proposed dataset includes 1,378 recent news articles in five languages focusing on the Ukraine-Russia War and climate change . the authors report evaluation results on state-of-the-art multilingual transformers and hierarchical zero-shot learning using LLMs at the level of a document, paragraph, and sentence . |
Text2Story Lusa: A Dataset for Narrative Analysis in European Portuguese News Articles (2024.lrec-main)
Copied to clipboard
Sérgio Nunes, Alípio Mario Jorge, Evelin Amorim, Hugo Sousa, António Leal, Purificação Moura Silvano, Inês Cantante, Ricardo Campos
| Challenge: | Access to annotated corpora with narrative elements is limited due to the lack of readily available datasets and copyright concerns. |
| Approach: | They developed a dataset that contains 357 news articles and 117 manually annotated articles with over 50 thousand individual annotations. |
| Outcome: | The proposed datasets are available in English and Portuguese and are based on 117 articles totaling over 50 thousand individual annotations. |
PartisanLens: A Multilingual Dataset of Hyperpartisan and Conspiratorial Immigration Narratives in European Media (2026.eacl-long)
Copied to clipboard
Michele Joshua Maggini, Paloma Piot, Anxo Pérez, Erik Bran Marino, Lúa Santamaría Montesinos, Ana Lisboa Cotovio, Marta Vázquez Abuín, Javier Parapar, Pablo Gamallo
| Challenge: | Existing methods for detecting hyperpartisan narratives and PRCTs are limited . hyperpartisan content promotes extreme views through one-sided, emotional language . |
| Approach: | They propose a multilingual dataset of 1617 hyperpartisan news headlines in Spanish, Italian, and Portuguese annotated in multiple political discourse aspects. |
| Outcome: | The proposed dataset is the first multilingual dataset of 1617 hyperpartisan headlines in Spanish, Italian, and Portuguese. |
Narratives at Conflict: Computational Analysis of News Framing in Multilingual Disinformation Campaigns (2024.acl-srw)
Copied to clipboard
| Challenge: | Existing methods for multilingual framing differ from those used in English-speaking world . framers often use loaded vocabularies to create political images or favor a particular point of view . |
| Approach: | They use eight years of Russian-backed disinformation campaigns to examine framing . they find that disinformation campaign consistently favors specific framers . |
| Outcome: | The proposed method underperforms and shows high disagreements in Russian-language articles . the proposed method is based on eight years of Russian-backed disinformation campaigns . |
Multilingual Multifaceted Understanding of Online News in Terms of Genre, Framing, and Persuasion Techniques (2023.acl-long)
Copied to clipboard
| Challenge: | a new dataset of news articles is presented that covers genre, framing, and persuasion techniques. |
| Approach: | They propose a multilingual multifacet dataset of news articles annotated for genre, framing and persuasion techniques. |
| Outcome: | The proposed dataset contains 1,612 news articles covering recent news on current topics of public interest in six European languages. |
A Study on Scaling Up Multilingual News Framing Analysis (2024.findings-naacl)
Copied to clipboard
| Challenge: | Existing studies on media framing have focused on English only data, leaving a gap in research concerning multilingual contexts. |
| Approach: | They propose to use crowd-sourced datasets to automate framing analysis by automating translation and annotation. |
| Outcome: | The proposed system improves on existing models in Bengali and Portuguese . the proposed system can train on a crowd-sourced dataset in 12 languages . |
text2story: A Python Toolkit to Extract and Visualize Story Components of Narrative Text (2024.lrec-main)
Copied to clipboard
| Challenge: | Story components, namely events, time, participants, and their relations, are present in narrative texts from different domains such as journalism, medicine, finance, and law. |
| Approach: | They propose to use an array of narrative extraction tools to extract narratives from text . the package contains an array and an experimental module for evaluation . |
| Outcome: | The text2story python supports the narrative extraction and visualization pipeline. |
Detecting Narrative Elements in Informational Text (2022.findings-naacl)
Copied to clipboard
| Challenge: | Recent work has focused on identifying narrative elements in personal stories texts, but this paper focuses on informational texts. |
| Approach: | They propose a novel NLP task for detecting narrative elements in raw text by adapting elements from the oral narrative theory of Labov and Waletzky and adding a new narrative element of their own. |
| Outcome: | The proposed scheme achieves an average F1 score of 0.77 and is better suited for informational texts than the oral narrative theory. |
A Dataset for Multi-lingual Epidemiological Event Extraction (2020.lrec-1)
Copied to clipboard
| Challenge: | Using the Web, we propose a corpus for information extraction and text classification. |
| Approach: | They propose to use a corpus for information extraction and natural language processing (NLP) tasks such as text classification. |
| Outcome: | The proposed corpus can be used for information extraction and natural language processing tasks such as text classification. |
Multi-Label and Multilingual News Framing Analysis (2020.acl-main)
Copied to clipboard
| Challenge: | Recent studies have focused on news framing in English, but few studies have explored how it can be extended to other languages and in multi-label settings. |
| Approach: | They propose a method that leverages dictionary and few annotations to detect frames from just the headline in a low-resource context. |
| Outcome: | The proposed method performs better than translating the entire headline to the source language . it can be scaled up to many languages, even those without existing translation technologies . |