Interannotator Agreement for Lexico-Semantic Annotation of a Corpus (2020.lrec-1)
Copied to clipboard
| Challenge: | a method for lexico-semantic annotation of the Basic Corpus of Polish Metaphors is described . the procedure is composed of three steps: deciding whether a particular occurrence of a word is asemantics or strictly grammatical. |
| Approach: | They propose a procedure for lexico-semantic annotation of the Basic Corpus of Polish Metaphor . procedure corrects morphosyntactic annotation of part of corpus that is automatically annotated . |
| Outcome: | The proposed procedure corrects the morphosyntactic annotation of part of the corpus . it is composed of three steps: deciding whether a word is asemantic or strictly grammatical . preliminary results show that the procedure is adequate for the task . |
Similar Papers
Annotation of metaphorical expressions in the Basic Corpus of Polish Metaphors (2022.lrec-1)
Copied to clipboard
| Challenge: | a corpus of Polish texts annotated with metaphorical expressions is composed of two parts of comparable size, selected from two subcorpora of the Polish National Corpus . |
| Approach: | They propose to use a procedure to annotate metaphorical expressions in Polish texts using two different subcorpora of the Polish National Corpus . they propose several features to classify metaphorical Expressions identified in texts. |
| Outcome: | The proposed procedure is based on the MIPVU procedure and focuses on neologistic derivatives that have metaphorical properties. |
Metaphor annotation for German (2022.lrec-1)
Copied to clipboard
| Challenge: | a corpus annotated for metaphors denotes entities or situations that are in some sense similar to the literal referent, but we believe it is of interest to research on metaphor in general. |
| Approach: | They present a German corpus annotated for metaphor in a project on register and propose to broaden the annotation to include metonymy. |
| Outcome: | The proposed corpus is compiled and annotated in a project on the interdependence of metaphors and register. |
Polish Corpus of Annotated Descriptions of Images (L18-1)
Copied to clipboard
| Challenge: | a new dataset of image descriptions is presented in Polish . the dataset is too small for training a sophisticated language-vision system. |
| Approach: | They propose to use a Polish dataset to analyze image descriptions . the descriptions are morphosyntactically analysed and annotated by human annotators . |
| Outcome: | The proposed model learns about the inter-modal correspondences between language and vision. |
Annotation and Automatic Classification of Aspectual Categories (P19-1)
Copied to clipboard
| Challenge: | Annotated resource for aspectual classification of German verb tokens in context. |
| Approach: | They present a resource for aspectual classification of German verb tokens in their clausal context. |
| Outcome: | The proposed resource is compared with previous work on German verb tokens using aspectual features compatible with the plurality of aspectual classifications. |
A corpus of metaphors as register markers (2023.findings-eacl)
Copied to clipboard
| Challenge: | Using corpus annotation, we show huge differences in metaphor usage between different registers and specific properties of registers. |
| Approach: | They present their work on corpus annotation for metaphor in germany . they focus on metaphors that can serve as register markers and be reliably indentified . |
| Outcome: | The proposed corpus annotations show huge differences in metaphor usage between different registers and specific properties of registers. |
An Empirical Evaluation of Annotation Practices in Corpora from Language Documentation (2020.lrec-1)
Copied to clipboard
| Challenge: | Language documentation projects have produced substantial amounts of primary data from a wide variety of endangered languages. |
| Approach: | They propose to use common annotation conventions in existing corpora to facilitate their future processing. |
| Outcome: | The proposed formats are based on the common ELAN and Toolbox formats and are used to facilitate their future processing. |
Manually Annotated Corpus of Polish Texts Published between 1830 and 1918 (L18-1)
Copied to clipboard
| Challenge: | a paper presents a manually annotated corpus of 625,000 tokens of Polish texts . the corpus provides three layers: transliteration, transcription and morphosyntactic annotation. |
| Approach: | The paper presents a manually annotated large historical corpus of Polish . the corpus provides three layers: transliteration, transcription and morphosyntactic annotation. |
| Outcome: | The corpus provides three layers: transliteration, transcription and morphosyntactic annotation. |
A Corpus of Non-Native Written English Annotated for Metaphor (N18-2)
Copied to clipboard
| Challenge: | Using argumentation-relevant metaphor predicts a holistic score of essay quality, we show . |
| Approach: | They present a corpus of argumentative essays annotated for metaphor by non-native speakers of English . they also examine the relationship between writing proficiency and metaphor use . |
| Outcome: | The proposed corpus is made publicly available and evaluated . it shows that metaphor is a significant predictor of a holistic score of essay quality . |
Annotating Arguments in a Corpus of Opinion Articles (2022.lrec-1)
Copied to clipboard
Gil Rocha, Luís Trigo, Henrique Lopes Cardoso, Rui Sousa-Silva, Paula Carvalho, Bruno Martins, Miguel Won
| Challenge: | Argument annotation is the process of exposing and justifying one's points of view, with the aim of conveying a logical reasoning through a set of semantically related propositions. |
| Approach: | They propose to use argumentative discourse units to annotate arguments in Portuguese using a multi-layered process to analyze the annotations produced. |
| Outcome: | The proposed model exploits the best practices identified in previous studies while fostering the potential use of the resulting annotated corpus for new purposes. |
A Short Survey on Sense-Annotated Corpora (2020.lrec-1)
Copied to clipboard
| Challenge: | Word Sense Disambiguation (WSD) is a key task in Natural Language Understanding. |
| Approach: | They propose to use sense-annotated corpora for supervised Word Sense Disambiguation. |
| Outcome: | The proposed methods have been compared with knowledge-based approaches and have shown to be more efficient when they are available. |