Challenge: Discourse information is difficult to represent and annotate, and corpora annotated under different frameworks vary considerably.
Approach: They propose to use automatic means to unify discourse structures and relations . they will also explore the application of the unified framework in multi-task learning and graphical models .
Outcome: The proposed method can be used in multi-task learning and graphical models.

Similar Papers

QUD-Based Annotation of Discourse Structure and Information Structure: Tool and Evaluation (L18-1)

Copied to clipboard

Challenge: a new annotation scheme and discourse-analytic method is developed for information structure annotation.
Approach: They propose a new annotation scheme and a discourse-analytic method based on Questions under Discussion . they introduce a tool which enables the analyst to semi-automatically segment texts and enhance them with QUDs .
Outcome: The proposed method achieves good inter-annotator scores and good agreement with discourse annotations.
A Survey of QUD Models for Discourse Processing (2025.naacl-long)

Copied to clipboard

Challenge: Question Under Discussion (QUD) is a linguistic analytic framework for explaining pragmatic phenomena and information structural analysis.
Approach: They propose to use Question Under Discussion (QUD) to model discourse units, such as sentences, as answers to some implicit or explicit questions.
Outcome: The proposed model is compared with RST, PDTB and SDRT . questions that may require further study are suggested.
Enhancing Discourse Dependency Parsing with Sentence Dependency Parsing: A Unified Generative Method Based on Code Representation (2024.findings-emnlp)

Copied to clipboard

Challenge: Existing annotation resources for Discourse Dependency Parsing tasks are limited due to their complexity and annotation schema differences.
Approach: They propose a code-based unified dependency parsing method that uses code to model dependency parses under different annotation schemas.
Outcome: The proposed method improves on two Chinese DDP tasks.
A Multi-layer Annotated Corpus of Argumentative Text: From Argument Schemes to Discourse Relations (L18-1)

Copied to clipboard

Challenge: Recent interest in Argumentation Mining has brought to the fore the need for corpora annotated with argument information, which can be used as training data.
Approach: They propose a set of guidelines for the annotation of argument schemes and a new annotation tool for the 'inferential' argument schemes.
Outcome: The proposed corpus includes 112 argumentative microtexts and a new annotation tool.
TIARA: A Tool for Annotating Discourse Relations and Sentence Reordering (2020.lrec-1)

Copied to clipboard

Challenge: Existing tools for discourse relations and sentence reordering are difficult to use and clutter the display.
Approach: They propose to use TIARA to simplify the annotation process by offering interactive visualisation, including coloured links, indentation, and dual-view.
Outcome: The proposed tool simplifies the annotation process and offers visualisations including coloured links, indentation, and dual-view.
Multilingual Extension of PDTB-Style Annotation: The Case of TED Multilingual Discourse Bank (L18-1)

Copied to clipboard

Challenge: Existing corpora enriched with discourse annotations are scarce but exist . TED-MDB is hoped to be a source of parallel data for contrastive linguistic analysis and language technology applications.
Approach: They propose a multilingual discourse treebank to provide a clear description of discourse structure and semantics in multiple languages.
Outcome: The proposed corpus provides a clearly described level of discourse structure and semantics in multiple languages.
SciDTB: Discourse Dependency TreeBank for Scientific Abstracts (P18-2)

Copied to clipboard

Challenge: Discourse relations are annotated on scientific articles.
Approach: They propose a domain-specific discourse treebank annotated on scientific articles . they use dependency trees to represent discourse structure, which is flexible and simplified .
Outcome: The proposed treebank is a benchmark for evaluating discourse dependency parsers.
Universal Anaphora: The First Three Years (2024.lrec-main)

Copied to clipboard

Challenge: Universal Anaphora initiative aims to push forward the state of the art in anaphora and anaphorism resolution by expanding the aspects of anaphonic interpretation which are or can be reliably annotated in an anagraphic corpora.
Approach: They propose to develop a standard for anaphoric annotations and a method for evaluating models that can carry out this type of interpretation.
Outcome: The Universal Anaphora initiative aims to push forward the state of the art in anaphora and anaphorism resolution by producing unified standards to annotate and encode annotations, delivering datasets encoded according to these standards, and developing methods for evaluating models that carry out this type of interpretation.
Joint Modeling of Entities and Discourse Relations for Coherence Assessment (2025.emnlp-main)

Copied to clipboard

Challenge: Existing work on coherence modeling focuses on entity features or discourse relation features, with little attention given to combining the two.
Approach: They propose two methods for jointly modeling entities and discourse relations for coherence assessment.
Outcome: The proposed methods significantly improve the performance of coherence models on three benchmark datasets.
Enriching a Lexicon of Discourse Connectives with Corpus-based Data (L18-1)

Copied to clipboard

Challenge: Existing annotation efforts for multiple languages have focused on discourse connectives, but we have limited it to the class of connectives marking contrast and the additional relations such connectives might convey.
Approach: They enrich a lexicon of italian COnnectives with real corpus data for connectives marking contrast relations in text.
Outcome: The proposed resource is a valuable tool for linguistic analyses of discourse relations and the training of a classifier for NLP applications.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations