Challenge: Infant-directed speech is often seen as a predictor for infants' speech processing abilities, for instance speech segmentation or word learning.
Approach: They examine the syntactic distribution, accentuation and prosodic phrasing of German verb forms and show that many verb forms are prime candidates for early segmentation.
Outcome: The findings suggest that infants ought to be able to extract verbs as early as nouns, given appropriate stimulus materials.

Similar Papers

Every Verb in Its Right Place? A Roadmap for Operationalizing Developmental Stages in the Acquisition of L2 German (2024.lrec-main)

Copied to clipboard

Challenge: Developmental stages are a linguistic concept claiming that language learning progresses in an ordered, step-like manner.
Approach: They propose to translate a linguistic specification into a computational procedure that can assign clauses to a developmental stage based on verb placement.
Outcome: The proposed system lacks a coherent linguistic specification of developmental stages . it also lacks the ability to translate the specification into a computational procedure based on verb placement.
GerEO: A Large-Scale Resource on the Syntactic Distribution of German Experiencer-Object Verbs (2022.lrec-1)

Copied to clipboard

Challenge: Psych verbs and their properties in multiple languages have ignited discussions among linguists for several decades . Psych-verbs are often considered syntactically deviant, although this has occasionally been called into question .
Approach: They propose to use a large-scale database of more than 10,000 examples for 64 verbs from a newspaper corpus annotated for several syntactic and semantic features relevant for their analysis.
Outcome: The proposed database contains 10,000 examples for 64 verbs from a newspaper corpus and includes syntactic construction, semantic stimulus type, and form of possible stimulus preposition.
Is Word Segmentation Child’s Play in All Languages? (P19-1)

Copied to clipboard

Challenge: Existing word learning strategies for infants are cross-linguistically robust . infants do not know which language(s) will be found in their environment at the beginning of development .
Approach: They propose to use 11 conceptually diverse algorithms to learn word-like units in infants . they propose to employ cross-linguistically robust algorithms that can be used by all infants.
Outcome: The proposed algorithms perform above chance on 8 different languages . the results show that some of the algorithms are cross-linguistically valid .
Annotation and Automatic Classification of Aspectual Categories (P19-1)

Copied to clipboard

Challenge: Annotated resource for aspectual classification of German verb tokens in context.
Approach: They present a resource for aspectual classification of German verb tokens in their clausal context.
Outcome: The proposed resource is compared with previous work on German verb tokens using aspectual features compatible with the plurality of aspectual classifications.
GeCoTagger: Annotation of German Verb Complements with Conditional Random Fields (L18-1)

Copied to clipboard

Challenge: Complement phrases are essential for constructing well-formed sentences in German.
Approach: They propose an algorithm which can identify and classify complement phrases of any German verb in any written sentence context.
Outcome: The proposed algorithm can identify and classify complement phrases of any German verb in any written sentence context.
Acquiring Verb Classes Through Bottom-Up Semantic Verb Clustering (L18-1)

Copied to clipboard

Challenge: Existing methods for creating verbal classifications are limited or non-existent in most languages . a range of automatic verb classification approaches have been proposed, but high-quality resources are needed .
Approach: They propose to use top-up semantic clustering to extract syntactic and semantic information from verbs in English, Polish and Croatian.
Outcome: The proposed classifications in English, Polish and Croatian are compared with other languages.
Learning to Understand Child-directed and Adult-directed Speech (2020.acl-main)

Copied to clipboard

Challenge: linguistic properties of child-directed speech differ from adult-directed in many ways . linguistic differences between CDS and ADS are retained, but the acoustic properties are similar.
Approach: They compare the task performance of models trained on adult-directed speech and child-directed language . they propose that CDS is optimized for learnability, but not for comprehension .
Outcome: The proposed model trains on adult-directed speech and child-directed language . the model generalizes better on the training register and on synthesized speech .
German Light Verb Constructions in Business Process Models (2022.lrec-1)

Copied to clipboard

Challenge: a resource of German light verb constructions is presented for graphical business process models . the language in BPM is worth to be studied for other purposes, authors say .
Approach: They present a resource of German light verb constructions extracted from business process models . they use textual labels to analyze the models and to infer the meaning of their texts .
Outcome: The proposed resource contains German light verb constructions extracted from business process models . the work is a step towards better automatic analysis of business process model models based on the proposed language .
Is Child-Directed Speech Effective Training Data for Language Models? (2024.emnlp-main)

Copied to clipboard

Challenge: High-performing language models are typically trained on hundreds of billions of words, but human learners use language fluently after far less training data.
Approach: They train GPT-2 and RoBERTa models on 29M words of English child-directed speech and a new matched, synthetic dataset.
Outcome: The proposed models show that child language input is not valuable for training language models.
Detecting Syntactic Change with Pre-trained Transformer Models (2023.findings-emnlp)

Copied to clipboard

Challenge: a fine-tuned BERT model can distinguish between text from the early 1800s and late 1900s . we use it to identify specific instances of syntactic change and specific words for which a new part of speech was introduced.
Approach: They propose to use a BERT-based model to find syntactic differences between English of the early 1800s and that of the late 1900s.
Outcome: The proposed model can distinguish between English of the early 1800s and that of the late 1900s using only syntactic information.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations