A Recorded Debating Dataset (L18-1)

Copied to clipboard

Challenge: Existing research in computational argumentation and debating technologies focuses on argumentation mining, but other tasks are being addressed as well.
Approach: They describe a dataset of debating speeches in English that is used for research . they use an automatic speech recognition system to produce a more "nLP-friendly" text .
Outcome: The proposed dataset contains 60 speeches on various controversial topics, each in five formats corresponding to different stages in production.

Similar Papers

FREDSum: A Dialogue Summarization Corpus for French Political Debates (2023.findings-emnlp)

Copied to clipboard

Challenge: Recent advances in deep learning have improved the performance of abstractive summarization systems.
Approach: They present a dataset of french political debates to enhance resources for multi-lingual dialogue summarization.
Outcome: The proposed dataset will be made publicly available for use by the research community.
IAM: A Comprehensive and Large-Scale Dataset for Integrated Argument Mining Tasks (2022.acl-long)

Copied to clipboard

Challenge: Argument mining (AM) is a computational process that is used to analyze information in a debating system.
Approach: They propose to use a large dataset to automate the manual process of debating . they propose to integrate claim extraction, stance classification and evidence extraction tasks .
Outcome: The proposed tasks can extract claims, stances, evidence and more from a large dataset . the proposed tasks are highly efficient and can be applied to argument mining tasks .
Advances in Debating Technologies: Building AI That Can Debate Humans (2021.acl-tutorials)

Copied to clipboard

Challenge: This tutorial focuses on Debating Technologies, a sub-field of computational argumentation defined as "computational technologies developed directly to enhance, support, and engage with human debating" the tutorial provides a holistic view of a debated system, and discusses practical applications and future challenges of debation technologies.
Approach: They present a tutorial on Debating Technologies, a sub-field of computational argumentation . they introduce Project Debater, which is the first AI system to debate human experts .
Outcome: The project Debater is the first AI system to debate human experts on complex topics.
Listening Comprehension over Argumentative Content (D18-1)

Copied to clipboard

Challenge: In argumentation domain, people are exposed directly to audio (or the video), without access to a written version.
Approach: They present a task for machine listening comprehension in the argumentation domain and a dataset in English.
Outcome: The proposed task is based on 200 speeches arguing for or against 50 controversial topics and uses baseline methods to address it.
Out of the Echo Chamber: Detecting Countering Debate Speeches (2020.acl-main)

Copied to clipboard

Challenge: Existing algorithms to detect articles that counter the arguments in debate speeches are unsuccessful, suggesting room for further research.
Approach: They propose a task to detect articles that counter the arguments made in debate speeches by annotating them from a dataset of 3,685 such speeches.
Outcome: The proposed algorithm can detect articles that counter the arguments made in debate speeches, and some are successful, but none are human-like.
ParlVote: A Corpus for Sentiment Analysis of Political Debates (2020.lrec-1)

Copied to clipboard

Challenge: Debate transcripts from the UK Parliament contain information about the positions taken by politicians towards important topics, but are difficult for humans to process manually.
Approach: They propose to use a linear classifier and a transformer word embedding model to classify sentiment polarity in debate speeches to evaluate sentiment analysis systems for the political domain.
Outcome: The proposed method performs better on the largest dataset and is more robust to other datasets.
100,000 Podcasts: A Spoken English Document Corpus (2020.coling-main)

Copied to clipboard

Challenge: Podcasts are a large and growing repository of spoken audio.
Approach: They propose to use podcasts as a resource for speech processing and linguistics . they use a corpus of 100,000 podcasts to study the complexity of the domain .
Outcome: The Spotify Podcast Dataset is the largest corpus of transcribed speech data . the dataset contains 60,000 hours of podcasts, with a range of genres and styles .
A Dataset of General-Purpose Rebuttal (D19-1)

Copied to clipboard

Challenge: a key element in argumentation is rebuttal, the ability to contest an argument by presenting a counter-argument.
Approach: They propose a method based on general rebuttal arguments to produce a critical response to a long argumentative text.
Outcome: The proposed method overcomes the need for topic-specific arguments to be provided . it allows creating responses beyond the scope of topics for which specific arguments are available .
The Discussion Tracker Corpus of Collaborative Argumentation (2020.lrec-1)

Copied to clipboard

Challenge: The Discussion Tracker corpus is an annotated dataset of transcripts of spoken, multi-party argumentation transcribed from 985 minutes of audio .
Approach: They analyze 29 multi-party arguments transcribed from 985 minutes of audio . they provide descriptive statistics and code for predicting each dimension separately.
Outcome: The Discussion Tracker corpus was collected in high school English classes and annotated for argument moves, specificity, specificities and collaboration dimensions.
Yes, we can! Mining Arguments in 50 Years of US Presidential Campaign Debates (P19-1)

Copied to clipboard

Challenge: Political debates are a natural application scenario for Argument Mining.
Approach: They propose an argument mining approach to political debates that uses argument components to annotate 39 political debate from the last 50 years of US presidential campaigns.
Outcome: The proposed approach outperforms baselines in argument mining over political debates.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations