Analyzing Middle High German Syntax with RDF and SPARQL (L18-1)

Copied to clipboard

Challenge: Using CoNLL-RDF and SPARQL Update, we analyze the diachronic changes of Middle High German syntax.
Approach: They propose a rule-based shallow parser and an enrichment pipeline grounded in CoNLL-RDF and SPARQL Update for parsing.
Outcome: The proposed pipeline is based on CoNLL-RDF and SPARQL Update for syntactic annotation and semantic enrichment of Middle High German.

Similar Papers

A Penn-style Treebank of Middle Low German (2020.lrec-1)

Copied to clipboard

Challenge: attestation for Middle Low German is rich, but its syntax remains relatively understudied.
Approach: They outline the issues involved in creating a Penn-style treebank of Middle Low German . they describe the background for the corpus and the process by which texts were selected .
Outcome: The proposed corpus will be a syntactically annotated treebank of Middle Low German . the proposed corpuse will be part of the Corpus of Historical Low German (CHLG) the proposed method will be used to generate strong empirical evidence for the language .
HiNTS: A Tagset for Middle Low German (L18-1)

Copied to clipboard

Challenge: a non-standardized language such as Middle Low German has special requirements for annotating part of speech and morphology.
Approach: They describe a tagset for annotating parts-of-speech and morphology in Middle Low German texts . they describe two special features of the tagse, and prove their usefulness .
Outcome: The proposed tagset can be used to annotate parts-of-speech and morphology in Middle Low German texts.
Universal Dependencies: Extensions for Modern and Historical German (2024.lrec-main)

Copied to clipboard

Challenge: a new UD treebank is being developed for Middle High German annotations . the annotation scheme is inconsistent with other treebanks for this period .
Approach: They propose to extend the UD scheme for modern and historical German by a range of tokens . they propose to use a treebank that is the first UD treebank for Middle High German .
Outcome: The proposed extensions relate in part to differences between arguments and modifiers . the proposed treebank is the first UD treebank for Middle High German .
Introducing a Parsed Corpus of Historical High German (2024.lrec-main)

Copied to clipboard

Challenge: outlines the development of the Indiana Parsed Corpus of (Historical) High German . outlines selection of texts, decisions on part-of-speech tags and other labels .
Approach: They propose to build a parsed German corpus that spans Germanic from 1050 to 1950 . they propose to use Penn-style treebanks to capture syntactic relationships between words .
Outcome: The proposed corpus spans Germanic languages from 1050 to 1950 and illustrative annotation issues unique to the language.
Tracing Syntactic Change in the Scientific Genre: Two Universal Dependency-parsed Diachronic Corpora of Scientific English and German (2022.lrec-1)

Copied to clipboard

Challenge: a recent study has focused on the syntactic development of scientific discourse in English and German.
Approach: They present two comparable diachronic corpora of scientific English and German from the Late Modern Period (17th c.–19th d.) annotated with Universal Dependencies.
Outcome: The presented corpora are comparable to existing studies on grammatical change in English and German . the results show that the pre-processing steps significantly improve parsing accuracy .
A database of German definitory contexts from selected web sources (L18-1)

Copied to clipboard

Challenge: a specialized web corpus and robust pattern-based extraction methods are used to detect definitory contexts.
Approach: They propose to use a web corpus and a database to detect definitory contexts . they describe an experimental setting and front-end for pattern-based definition extraction .
Outcome: The proposed method is based on a web corpus and a robust pattern-based extraction method.
Syntax in End-to-End Natural Language Processing (2021.emnlp-tutorials)

Copied to clipboard

Challenge: tutorial focuses on syntactic parsing and syntax in end-to-end natural language processing (NLP) tasks.
Approach: tutorial will introduce syntactic parsing and the role of syntax in end-to-end natural language processing (NLP) tasks.
Outcome: This tutorial will introduce the background and the latest progress of syntactic parsing and SRL/NMT.
Annotation and Automatic Classification of Aspectual Categories (P19-1)

Copied to clipboard

Challenge: Annotated resource for aspectual classification of German verb tokens in context.
Approach: They present a resource for aspectual classification of German verb tokens in their clausal context.
Outcome: The proposed resource is compared with previous work on German verb tokens using aspectual features compatible with the plurality of aspectual classifications.
The Low Saxon LSDC Dataset at Universal Dependencies (2024.lrec-main)

Copied to clipboard

Challenge: Low Saxon is a low-resource language that lacks a common standard . dialectal variation in morphological categories can cause problems .
Approach: They extend the Low Saxon Universal Dependencies dataset to include 8 of the 9 major dialects.
Outcome: The proposed dataset covers the last 200 years and 8 of the 9 major dialects.
GerEO: A Large-Scale Resource on the Syntactic Distribution of German Experiencer-Object Verbs (2022.lrec-1)

Copied to clipboard

Challenge: Psych verbs and their properties in multiple languages have ignited discussions among linguists for several decades . Psych-verbs are often considered syntactically deviant, although this has occasionally been called into question .
Approach: They propose to use a large-scale database of more than 10,000 examples for 64 verbs from a newspaper corpus annotated for several syntactic and semantic features relevant for their analysis.
Outcome: The proposed database contains 10,000 examples for 64 verbs from a newspaper corpus and includes syntactic construction, semantic stimulus type, and form of possible stimulus preposition.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations