Challenge: lexical forms and prosodic characteristics of short feedback tokens are not indicative of their communicative function.
Approach: They propose to annotate short feedback tokens with a lexical annotation scheme . they find that feedback functions have distinguishable prosodic characteristics .
Outcome: The proposed annotations show that lexical forms alone are not indicative of the communicative function.

Similar Papers

Developing a Benchmark for Pronunciation Feedback: Creation of a Phonemically Annotated Speech Corpus of isiZulu Language Learner Speech (2024.lrec-main)

Copied to clipboard

Challenge: Existing corpora for computer-assisted pronunciation training (CAPT) do not apply well to research in pronunciation feedback.
Approach: They propose to create a corpus of isiZulu language learner speech that has been annotated for phoneme errors and suprasegmental errors in tone.
Outcome: The proposed corpus is comprised of gold standard recordings from isiZulu teachers and recordings from students that have been annotated for pronunciation errors.
Creating Corpora for Research in Feedback Comment Generation (2020.lrec-1)

Copied to clipboard

Challenge: Existing corpus of learner corpora with feedback comments is limited due to the lack of public access to this task.
Approach: They describe two corpora that have been manually annotated with feedback comments . they describe how the principle and guidelines for feedback comment annotation work .
Outcome: The proposed corpus is available on the web and will facilitate research in feedback comment generation.
Finding Common Ground: Annotating and Predicting Common Ground in Spoken Conversations (2023.findings-emnlp)

Copied to clipboard

Challenge: Creating and updating common ground (CG) between interlocutors is the key to a successful conversation.
Approach: They propose a new annotation and corpus to capture common ground in human communication . they then conduct experiments to extract propositions from dialog and track their status in common ground from the perspective of each speaker .
Outcome: The proposed corpus captures common ground from the perspective of two speakers in a dialog.
An Evaluation Dataset for Identifying Communicative Functions of Sentences in English Scholarly Papers (2020.lrec-1)

Copied to clipboard

Challenge: Formulaic expressions are used by authors of scientific papers because they convey specific communicative functions in the rhetorical structure of papers.
Approach: They created a manually annotated dataset to detect formulaic expressions in sentences using a seed list of labelled formulaic words.
Outcome: The proposed dataset can detect communicative functions in sentences using a seed list of labelled expressions from scholarly papers in the ACL Anthology.
Annotating Interruption in Dyadic Human Interaction (2022.lrec-1)

Copied to clipboard

Challenge: Existing interruption and turn switch classification methods are not yet available.
Approach: They propose a new interruption annotation schema that integrates existing interruption and turn switch classification methods to annotate different types of interruptions.
Outcome: The proposed method can distinguish smooth turn exchange, backchannel and interruption (including interruption types) and to annotate dyadic conversation.
A Manually Annotated Resource for the Investigation of Nasal Grunts (2020.lrec-1)

Copied to clipboard

Challenge: acoustic annotation of nasal grunts is described in the whole CID corpus of the french language . acculturation of non-lexical conversational sounds has been debated for a long time .
Approach: They propose an annotation framework for nasal grunts of the whole French CID corpus . they characterise acoustic cues and visual cue conventions followed for the annotation .
Outcome: The proposed framework is based on the entire French CID corpus.
Dataset and Baseline for Automatic Student Feedback Analysis (2022.lrec-1)

Copied to clipboard

Challenge: Currently, student feedback is collected manually, but it does not indicate the student's opinion on different aspects of the teaching/learning process.
Approach: They propose to annotate student feedback corpus which contains 3000 instances . they propose a hierarchical taxonomy for aspect categorization, which covers all areas .
Outcome: The proposed model can be used for aspects analysis, document level sentiment analysis and document level analysis.
Annotation and Automatic Classification of Aspectual Categories (P19-1)

Copied to clipboard

Challenge: Annotated resource for aspectual classification of German verb tokens in context.
Approach: They present a resource for aspectual classification of German verb tokens in their clausal context.
Outcome: The proposed resource is compared with previous work on German verb tokens using aspectual features compatible with the plurality of aspectual classifications.
Discovering the Functions of Language in Online Forums (D19-55)

Copied to clipboard

Challenge: a vast amount of work has been dedicated to speech act categorization for characterizing discourses . lack of formalism and diversity of taxonomies make it difficult to compare different annotated datasets.
Approach: They propose a semi-supervised framework for predicting the functions of Reddit comments . they propose to use the framework to analyze online forum conversations .
Outcome: The proposed framework can predict functions of Reddit comments and 165K comments.
Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning (2026.acl-long)

Copied to clipboard

Challenge: Prior work on predicting backchannel timing has focused on lexical form and prosody, but the relationship between lexico-prosodic form and meaning remains underexplored.
Approach: They propose a framework for fine-tuning large language models on dialogue transcripts to derive rich contextual representations; and a joint embedding space for dialogue contexts and backchannel realizations.
Outcome: The proposed framework improves context-backchannel retrieval and human perception is more sensitive to extended conversational context and embeddings align more closely with human judgments than raw WavLM features.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations