An Assessment of Explicit Inter- and Intra-sentential Discourse Connectives in Turkish Discourse Bank (L18-1)
Copied to clipboard
| Challenge: | Discourse parsing is a challenging task for NLP. |
| Approach: | They propose to add a new set of explicit intra-sentential connectives to Turkish Discourse Bank 1.1 . they propose to evaluate the converb sense annotations and compare them to other Turkish corpus . |
| Outcome: | The proposed annotations show that the subordinators tend to select certain senses not selected by explicit inter- and intra-sentential discourse connectives in the data. |
Similar Papers
Enriching a Lexicon of Discourse Connectives with Corpus-based Data (L18-1)
Copied to clipboard
| Challenge: | Existing annotation efforts for multiple languages have focused on discourse connectives, but we have limited it to the class of connectives marking contrast and the additional relations such connectives might convey. |
| Approach: | They enrich a lexicon of italian COnnectives with real corpus data for connectives marking contrast relations in text. |
| Outcome: | The proposed resource is a valuable tool for linguistic analyses of discourse relations and the training of a classifier for NLP applications. |
Discourse Sense Flows: Modelling the Rhetorical Style of Documents across Various Domains (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Recent work on shallow discourse parsing has given renewed attention to the role of discourse relation signals, in particular explicit connectives and alternative lexicalizations. |
| Approach: | They propose a model for extracting and classifying discourse relation signals from the Penn Discourse Treebank v3 corpus and introduce a new way of modeling rhetorical style by the linear order of coherence relations. |
| Outcome: | The proposed models are based on the Penn Discourse Treebank v3 corpus and employ n-gram patterns to predict genre/domain discrimination. |
Towards Identifying Alternative-Lexicalization Signals of Discourse Relations (2022.coling-1)
Copied to clipboard
| Challenge: | Existing shallow discourse parsing methods have been limited to identifying relations signaled by a discourse connective and those without a signal. |
| Approach: | They propose to identify relations signalled by a discourse connective and those without . they compare a pattern-based approach and a sequence labeling model . |
| Outcome: | The proposed approach is based on a pattern-based approach and a sequence labeling model. |
Implicit Discourse Relation Classification: We Need to Talk about Evaluation (2020.acl-main)
Copied to clipboard
| Challenge: | Lack of consistency in preprocessing and evaluation poses challenges to fair comparison of results in literature. |
| Approach: | They propose an improved evaluation protocol for implicit relation classification on PDTB 2.0 . they report strong baseline results from pretrained sentence encoders . |
| Outcome: | The proposed evaluation protocol improves the existing framework and provides strong baseline results. |
Multilingual Extension of PDTB-Style Annotation: The Case of TED Multilingual Discourse Bank (L18-1)
Copied to clipboard
| Challenge: | Existing corpora enriched with discourse annotations are scarce but exist . TED-MDB is hoped to be a source of parallel data for contrastive linguistic analysis and language technology applications. |
| Approach: | They propose a multilingual discourse treebank to provide a clear description of discourse structure and semantics in multiple languages. |
| Outcome: | The proposed corpus provides a clearly described level of discourse structure and semantics in multiple languages. |
A Gold Standard Dependency Treebank for Turkish (2020.lrec-1)
Copied to clipboard
| Challenge: | Currently, Turkish treebanks are limited due to the limited number of annotated sentences in the domains of Wikipedia and ITU Web Treebanks. |
| Approach: | They propose to annotate Turkish web and Wikipedia sentences for segmentation, morphology, part-of-speech and dependency relations using tagsets and a Wikipedia section. |
| Outcome: | The proposed treebank is the largest publicly available morpho-syntactic treebank in terms of word count and has a dedicated Wikipedia section. |
Persian Discourse Treebank and coreference corpus (L18-1)
Copied to clipboard
| Challenge: | Currently, we are adding a new document-level discourse annotation to our new corpus. |
| Approach: | They propose to build a Persian discourse treebank and a comprehensive Persian coreference corpus based on discourse analysis and coreference resolution. |
| Outcome: | The proposed corpus includes 30000 individual sentences with morphological, syntactic and semantic labels and nearly half a million tokens. |
Clarifying Underspecified Discourse Relations in Instructional Texts (2025.findings-acl)
Copied to clipboard
| Challenge: | Discourse relations can be optionally realized through explicit connectives such as “but” and “while”. |
| Approach: | They build a corpus of 4,274 text revisions in which a connective was explicitly inserted . they collect plausibility annotations on other connectives to check whether they represent suitable alternatives . |
| Outcome: | The proposed model predicts plausibility of individual connectives with up to 66% accuracy, but is not reliable when multiple relations are plausible. |
TRopBank: Turkish PropBank V2.0 (2020.lrec-1)
Copied to clipboard
| Challenge: | PropBank is a hand-annotated corpus of propositions used to obtain predicate-argument information of a language. |
| Approach: | They present TRopBank "Turkish PropBank v2.0" which is a hand-annotated corpus of propositions . it is used to obtain the predicate-argument information of a language . |
| Outcome: | The proposed annotations provide the predicate-argument information of a language . the proposed annotation is based on the annotations of 17.673 verbs in Turkish . |
Announcing the Prague Discourse Treebank 3.0 (2024.lrec-main)
Copied to clipboard
| Challenge: | PDiT 3.0 contains 21,662 discourse relations (plus 445 list relations) in 49 thousand sentences. |
| Approach: | They present the Prague Discourse Treebank 3.0, a new version of the annotation of discourse relations marked by primary and secondary discourse connectives in the Prague Dependency Treebank. |
| Outcome: | The new version of the PDiT 3.0 brings a largely revised annotation of discourse relations and achieves consistency with a Lexicon of Czech Discourse Connectives (CzeDLex) and sense taxonomy. |