The Distribution and Prosodic Realization of Verb Forms in German Infant-Directed Speech (L18-1)
Copied to clipboard
| Challenge: | Infant-directed speech is often seen as a predictor for infants' speech processing abilities, for instance speech segmentation or word learning. |
| Approach: | They examine the syntactic distribution, accentuation and prosodic phrasing of German verb forms and show that many verb forms are prime candidates for early segmentation. |
| Outcome: | The findings suggest that infants ought to be able to extract verbs as early as nouns, given appropriate stimulus materials. |
Similar Papers
Every Verb in Its Right Place? A Roadmap for Operationalizing Developmental Stages in the Acquisition of L2 German (2024.lrec-main)
Copied to clipboard
| Challenge: | Developmental stages are a linguistic concept claiming that language learning progresses in an ordered, step-like manner. |
| Approach: | They propose to translate a linguistic specification into a computational procedure that can assign clauses to a developmental stage based on verb placement. |
| Outcome: | The proposed system lacks a coherent linguistic specification of developmental stages . it also lacks the ability to translate the specification into a computational procedure based on verb placement. |
GerEO: A Large-Scale Resource on the Syntactic Distribution of German Experiencer-Object Verbs (2022.lrec-1)
Copied to clipboard
| Challenge: | Psych verbs and their properties in multiple languages have ignited discussions among linguists for several decades . Psych-verbs are often considered syntactically deviant, although this has occasionally been called into question . |
| Approach: | They propose to use a large-scale database of more than 10,000 examples for 64 verbs from a newspaper corpus annotated for several syntactic and semantic features relevant for their analysis. |
| Outcome: | The proposed database contains 10,000 examples for 64 verbs from a newspaper corpus and includes syntactic construction, semantic stimulus type, and form of possible stimulus preposition. |
Is Word Segmentation Child’s Play in All Languages? (P19-1)
Copied to clipboard
| Challenge: | Existing word learning strategies for infants are cross-linguistically robust . infants do not know which language(s) will be found in their environment at the beginning of development . |
| Approach: | They propose to use 11 conceptually diverse algorithms to learn word-like units in infants . they propose to employ cross-linguistically robust algorithms that can be used by all infants. |
| Outcome: | The proposed algorithms perform above chance on 8 different languages . the results show that some of the algorithms are cross-linguistically valid . |
Annotation and Automatic Classification of Aspectual Categories (P19-1)
Copied to clipboard
| Challenge: | Annotated resource for aspectual classification of German verb tokens in context. |
| Approach: | They present a resource for aspectual classification of German verb tokens in their clausal context. |
| Outcome: | The proposed resource is compared with previous work on German verb tokens using aspectual features compatible with the plurality of aspectual classifications. |
GeCoTagger: Annotation of German Verb Complements with Conditional Random Fields (L18-1)
Copied to clipboard
| Challenge: | Complement phrases are essential for constructing well-formed sentences in German. |
| Approach: | They propose an algorithm which can identify and classify complement phrases of any German verb in any written sentence context. |
| Outcome: | The proposed algorithm can identify and classify complement phrases of any German verb in any written sentence context. |
Acquiring Verb Classes Through Bottom-Up Semantic Verb Clustering (L18-1)
Copied to clipboard
| Challenge: | Existing methods for creating verbal classifications are limited or non-existent in most languages . a range of automatic verb classification approaches have been proposed, but high-quality resources are needed . |
| Approach: | They propose to use top-up semantic clustering to extract syntactic and semantic information from verbs in English, Polish and Croatian. |
| Outcome: | The proposed classifications in English, Polish and Croatian are compared with other languages. |
Learning to Understand Child-directed and Adult-directed Speech (2020.acl-main)
Copied to clipboard
| Challenge: | linguistic properties of child-directed speech differ from adult-directed in many ways . linguistic differences between CDS and ADS are retained, but the acoustic properties are similar. |
| Approach: | They compare the task performance of models trained on adult-directed speech and child-directed language . they propose that CDS is optimized for learnability, but not for comprehension . |
| Outcome: | The proposed model trains on adult-directed speech and child-directed language . the model generalizes better on the training register and on synthesized speech . |
German Light Verb Constructions in Business Process Models (2022.lrec-1)
Copied to clipboard
| Challenge: | a resource of German light verb constructions is presented for graphical business process models . the language in BPM is worth to be studied for other purposes, authors say . |
| Approach: | They present a resource of German light verb constructions extracted from business process models . they use textual labels to analyze the models and to infer the meaning of their texts . |
| Outcome: | The proposed resource contains German light verb constructions extracted from business process models . the work is a step towards better automatic analysis of business process model models based on the proposed language . |
Is Child-Directed Speech Effective Training Data for Language Models? (2024.emnlp-main)
Copied to clipboard
| Challenge: | High-performing language models are typically trained on hundreds of billions of words, but human learners use language fluently after far less training data. |
| Approach: | They train GPT-2 and RoBERTa models on 29M words of English child-directed speech and a new matched, synthetic dataset. |
| Outcome: | The proposed models show that child language input is not valuable for training language models. |
Detecting Syntactic Change with Pre-trained Transformer Models (2023.findings-emnlp)
Copied to clipboard
| Challenge: | a fine-tuned BERT model can distinguish between text from the early 1800s and late 1900s . we use it to identify specific instances of syntactic change and specific words for which a new part of speech was introduced. |
| Approach: | They propose to use a BERT-based model to find syntactic differences between English of the early 1800s and that of the late 1900s. |
| Outcome: | The proposed model can distinguish between English of the early 1800s and that of the late 1900s using only syntactic information. |