| Challenge: | Using CoNLL-RDF and SPARQL Update, we analyze the diachronic changes of Middle High German syntax. |
| Approach: | They propose a rule-based shallow parser and an enrichment pipeline grounded in CoNLL-RDF and SPARQL Update for parsing. |
| Outcome: | The proposed pipeline is based on CoNLL-RDF and SPARQL Update for syntactic annotation and semantic enrichment of Middle High German. |
Similar Papers
A Penn-style Treebank of Middle Low German (2020.lrec-1)
Copied to clipboard
| Challenge: | attestation for Middle Low German is rich, but its syntax remains relatively understudied. |
| Approach: | They outline the issues involved in creating a Penn-style treebank of Middle Low German . they describe the background for the corpus and the process by which texts were selected . |
| Outcome: | The proposed corpus will be a syntactically annotated treebank of Middle Low German . the proposed corpuse will be part of the Corpus of Historical Low German (CHLG) the proposed method will be used to generate strong empirical evidence for the language . |
HiNTS: A Tagset for Middle Low German (L18-1)
Copied to clipboard
| Challenge: | a non-standardized language such as Middle Low German has special requirements for annotating part of speech and morphology. |
| Approach: | They describe a tagset for annotating parts-of-speech and morphology in Middle Low German texts . they describe two special features of the tagse, and prove their usefulness . |
| Outcome: | The proposed tagset can be used to annotate parts-of-speech and morphology in Middle Low German texts. |
Universal Dependencies: Extensions for Modern and Historical German (2024.lrec-main)
Copied to clipboard
| Challenge: | a new UD treebank is being developed for Middle High German annotations . the annotation scheme is inconsistent with other treebanks for this period . |
| Approach: | They propose to extend the UD scheme for modern and historical German by a range of tokens . they propose to use a treebank that is the first UD treebank for Middle High German . |
| Outcome: | The proposed extensions relate in part to differences between arguments and modifiers . the proposed treebank is the first UD treebank for Middle High German . |
Introducing a Parsed Corpus of Historical High German (2024.lrec-main)
Copied to clipboard
| Challenge: | outlines the development of the Indiana Parsed Corpus of (Historical) High German . outlines selection of texts, decisions on part-of-speech tags and other labels . |
| Approach: | They propose to build a parsed German corpus that spans Germanic from 1050 to 1950 . they propose to use Penn-style treebanks to capture syntactic relationships between words . |
| Outcome: | The proposed corpus spans Germanic languages from 1050 to 1950 and illustrative annotation issues unique to the language. |
Tracing Syntactic Change in the Scientific Genre: Two Universal Dependency-parsed Diachronic Corpora of Scientific English and German (2022.lrec-1)
Copied to clipboard
| Challenge: | a recent study has focused on the syntactic development of scientific discourse in English and German. |
| Approach: | They present two comparable diachronic corpora of scientific English and German from the Late Modern Period (17th c.–19th d.) annotated with Universal Dependencies. |
| Outcome: | The presented corpora are comparable to existing studies on grammatical change in English and German . the results show that the pre-processing steps significantly improve parsing accuracy . |
A database of German definitory contexts from selected web sources (L18-1)
Copied to clipboard
| Challenge: | a specialized web corpus and robust pattern-based extraction methods are used to detect definitory contexts. |
| Approach: | They propose to use a web corpus and a database to detect definitory contexts . they describe an experimental setting and front-end for pattern-based definition extraction . |
| Outcome: | The proposed method is based on a web corpus and a robust pattern-based extraction method. |
Syntax in End-to-End Natural Language Processing (2021.emnlp-tutorials)
Copied to clipboard
| Challenge: | tutorial focuses on syntactic parsing and syntax in end-to-end natural language processing (NLP) tasks. |
| Approach: | tutorial will introduce syntactic parsing and the role of syntax in end-to-end natural language processing (NLP) tasks. |
| Outcome: | This tutorial will introduce the background and the latest progress of syntactic parsing and SRL/NMT. |
Annotation and Automatic Classification of Aspectual Categories (P19-1)
Copied to clipboard
| Challenge: | Annotated resource for aspectual classification of German verb tokens in context. |
| Approach: | They present a resource for aspectual classification of German verb tokens in their clausal context. |
| Outcome: | The proposed resource is compared with previous work on German verb tokens using aspectual features compatible with the plurality of aspectual classifications. |
The Low Saxon LSDC Dataset at Universal Dependencies (2024.lrec-main)
Copied to clipboard
| Challenge: | Low Saxon is a low-resource language that lacks a common standard . dialectal variation in morphological categories can cause problems . |
| Approach: | They extend the Low Saxon Universal Dependencies dataset to include 8 of the 9 major dialects. |
| Outcome: | The proposed dataset covers the last 200 years and 8 of the 9 major dialects. |
GerEO: A Large-Scale Resource on the Syntactic Distribution of German Experiencer-Object Verbs (2022.lrec-1)
Copied to clipboard
| Challenge: | Psych verbs and their properties in multiple languages have ignited discussions among linguists for several decades . Psych-verbs are often considered syntactically deviant, although this has occasionally been called into question . |
| Approach: | They propose to use a large-scale database of more than 10,000 examples for 64 verbs from a newspaper corpus annotated for several syntactic and semantic features relevant for their analysis. |
| Outcome: | The proposed database contains 10,000 examples for 64 verbs from a newspaper corpus and includes syntactic construction, semantic stimulus type, and form of possible stimulus preposition. |