| Challenge: | Existing resources for recognizing attributions in context are limited in size and completeness. |
| Approach: | They propose to use the largest and most complete attribution relations corpus to date . they propose to create sophisticated end-to-end solutions for attribution extraction . |
| Outcome: | The political news attribution relations corpus 2016 is the largest and most complete attribution relations corpuse to date. |
Similar Papers
An Environment for Relational Annotation of Political Debates (P19-3)
Copied to clipboard
| Challenge: | Scalable text analysis techniques can open corpora to new questions in computational social sciences and digital humanities. |
| Approach: | They describe a tool that allows annotating newspaper text with rich information about claims (demands) raised by politicians and other actors. |
| Outcome: | The MARDY tool realizes the complete workflow necessary for annotating a large newspaper text collection with rich information about claims (demands) raised by politicians and other actors. |
Out of the Mouths of MPs: Speaker Attribution in Parliamentary Debates (2024.lrec-main)
Copied to clipboard
| Challenge: | Identifying who says what to whom is an essential prerequisite for analysing human communication. |
| Approach: | They propose a new corpus for speaker attribution in german parliamentary debates . the data includes more than 7,700 manually annotated events of speech, thought and writing . they then apply their model to predict speech events in 20 years of debates and investigate the use of factives in the rhetoric of MPs. |
| Outcome: | The proposed model predicts speech events in 20 years of debates and investigates the use of factives in the rhetoric of MPs. |
DirectQuote: A Dataset for Direct Quotation Extraction and Attribution in News Articles (2022.lrec-1)
Copied to clipboard
| Challenge: | Existing methods to extract and attribute quotations from news data are difficult and require a lot of effort. |
| Approach: | They propose a corpus of 19,760 paragraphs and 10,279 direct quotations manually annotated from online news media. |
| Outcome: | The proposed corpus contains 19,760 paragraphs and 10,279 direct quotations manually annotated from online news media. |
Who Blames or Endorses Whom? Entity-to-Entity Directed Sentiment Extraction in News Text (2021.findings-acl)
Copied to clipboard
| Challenge: | Existing methods for sentiment analysis do not consider direction of sentiments between political entities. |
| Approach: | They propose a novel task of identifying directed sentiment relationship between political entities from a given news document. |
| Outcome: | The proposed method is useful for social science research questions in the 2016 election and COVID-19. |
The Pragmatics behind Politics: Modelling Metaphor, Framing and Emotion in Political Discourse (2020.findings-emnlp)
Copied to clipboard
| Challenge: | Existing computational models of political discourse do not incorporate metaphor and emotion in their functions. |
| Approach: | They propose to combine metaphor, emotion and political rhetoric to model political discourse . they show that they advance in three tasks: predicting political perspective of news articles, party affiliation of politicians and framing of policy issues. |
| Outcome: | The proposed models improve political discourse prediction, party affiliation and framing of policy issues. |
Discovering Biased News Articles Leveraging Multiple Human Annotations (2020.lrec-1)
Copied to clipboard
| Challenge: | Political propaganda and one-sided views can be found in the news and can cause distrust in media. |
| Approach: | They propose to annotate politically biased news articles by an algorithm annotated by domain experts and crowd workers and to compare them to crowd workers. |
| Outcome: | The proposed method compares domain experts to crowd workers and shows that bias can be detected automatically. |
PolitiCause: An Annotation Scheme and Corpus for Causality in Political Texts (2024.lrec-main)
Copied to clipboard
| Challenge: | PolitiCAUSE is a new corpus of political texts annotated for causality . it provides a detailed and robust annotation scheme for analyzing causal information . |
| Approach: | They propose a new corpus of political texts annotated for causality . they provide a detailed and robust annotation scheme for annotating causal information . |
| Outcome: | The proposed method achieves a moderate performance on the dataset, with a MCC score of 0.62. |
We Can Detect Your Bias: Predicting the Political Ideology of News Articles (2020.emnlp-main)
Copied to clipboard
| Challenge: | a new study examines the role of media in predicting political ideology or bias in news articles . systematic exposure to bias in the news can foster intolerance and ideological segregation . |
| Approach: | They propose an adversarial media adaptation and a specially adapted triplet loss for predicting political ideology in news articles. |
| Outcome: | The proposed model improves over state-of-the-art models in this challenging setup. |
Crossing the Aisle: Unveiling Partisan and Counter-Partisan Events in News Reporting (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Prior work in NLP has only studied media bias via linguistic style and word usage. |
| Approach: | They annotate a dataset containing 8,511 (counter-)partisan event annotations in 304 news articles from ideologically diverse media outlets. |
| Outcome: | The proposed dataset contains 8,511 (counter-)partisan event annotations in 304 news articles from ideologically diverse media outlets. |
Dataset of Quotation Attribution in German News Articles (2024.lrec-main)
Copied to clipboard
| Challenge: | Lack of annotated data for quotation attribution in news articles severely limits the quality and usability of possible systems. |
| Approach: | They propose a dataset for quotation attribution in German news articles using WIKINEWS and manually annotated quotes from 1000 articles. |
| Outcome: | The proposed dataset provides curated, high-quality annotations across 1000 documents (250,000 tokens) in a fine-grained annotation schema enabling various downstream uses for the dataset. |