Annotation of Communicative Functions of Short Feedback Tokens in Switchboard (2022.lrec-1)
Copied to clipboard
| Challenge: | lexical forms and prosodic characteristics of short feedback tokens are not indicative of their communicative function. |
| Approach: | They propose to annotate short feedback tokens with a lexical annotation scheme . they find that feedback functions have distinguishable prosodic characteristics . |
| Outcome: | The proposed annotations show that lexical forms alone are not indicative of the communicative function. |
Similar Papers
Developing a Benchmark for Pronunciation Feedback: Creation of a Phonemically Annotated Speech Corpus of isiZulu Language Learner Speech (2024.lrec-main)
Copied to clipboard
Alexandra O’Neil, Nils Hjortnaes, Francis Tyers, Zinhle Nkosi, Thulile Ndlovu, Zanele Mlondo, Ngami Phumzile Pewa
| Challenge: | Existing corpora for computer-assisted pronunciation training (CAPT) do not apply well to research in pronunciation feedback. |
| Approach: | They propose to create a corpus of isiZulu language learner speech that has been annotated for phoneme errors and suprasegmental errors in tone. |
| Outcome: | The proposed corpus is comprised of gold standard recordings from isiZulu teachers and recordings from students that have been annotated for pronunciation errors. |
Creating Corpora for Research in Feedback Comment Generation (2020.lrec-1)
Copied to clipboard
| Challenge: | Existing corpus of learner corpora with feedback comments is limited due to the lack of public access to this task. |
| Approach: | They describe two corpora that have been manually annotated with feedback comments . they describe how the principle and guidelines for feedback comment annotation work . |
| Outcome: | The proposed corpus is available on the web and will facilitate research in feedback comment generation. |
Finding Common Ground: Annotating and Predicting Common Ground in Spoken Conversations (2023.findings-emnlp)
Copied to clipboard
| Challenge: | Creating and updating common ground (CG) between interlocutors is the key to a successful conversation. |
| Approach: | They propose a new annotation and corpus to capture common ground in human communication . they then conduct experiments to extract propositions from dialog and track their status in common ground from the perspective of each speaker . |
| Outcome: | The proposed corpus captures common ground from the perspective of two speakers in a dialog. |
An Evaluation Dataset for Identifying Communicative Functions of Sentences in English Scholarly Papers (2020.lrec-1)
Copied to clipboard
| Challenge: | Formulaic expressions are used by authors of scientific papers because they convey specific communicative functions in the rhetorical structure of papers. |
| Approach: | They created a manually annotated dataset to detect formulaic expressions in sentences using a seed list of labelled formulaic words. |
| Outcome: | The proposed dataset can detect communicative functions in sentences using a seed list of labelled expressions from scholarly papers in the ACL Anthology. |
Annotating Interruption in Dyadic Human Interaction (2022.lrec-1)
Copied to clipboard
| Challenge: | Existing interruption and turn switch classification methods are not yet available. |
| Approach: | They propose a new interruption annotation schema that integrates existing interruption and turn switch classification methods to annotate different types of interruptions. |
| Outcome: | The proposed method can distinguish smooth turn exchange, backchannel and interruption (including interruption types) and to annotate dyadic conversation. |
A Manually Annotated Resource for the Investigation of Nasal Grunts (2020.lrec-1)
Copied to clipboard
| Challenge: | acoustic annotation of nasal grunts is described in the whole CID corpus of the french language . acculturation of non-lexical conversational sounds has been debated for a long time . |
| Approach: | They propose an annotation framework for nasal grunts of the whole French CID corpus . they characterise acoustic cues and visual cue conventions followed for the annotation . |
| Outcome: | The proposed framework is based on the entire French CID corpus. |
Dataset and Baseline for Automatic Student Feedback Analysis (2022.lrec-1)
Copied to clipboard
| Challenge: | Currently, student feedback is collected manually, but it does not indicate the student's opinion on different aspects of the teaching/learning process. |
| Approach: | They propose to annotate student feedback corpus which contains 3000 instances . they propose a hierarchical taxonomy for aspect categorization, which covers all areas . |
| Outcome: | The proposed model can be used for aspects analysis, document level sentiment analysis and document level analysis. |
Annotation and Automatic Classification of Aspectual Categories (P19-1)
Copied to clipboard
| Challenge: | Annotated resource for aspectual classification of German verb tokens in context. |
| Approach: | They present a resource for aspectual classification of German verb tokens in their clausal context. |
| Outcome: | The proposed resource is compared with previous work on German verb tokens using aspectual features compatible with the plurality of aspectual classifications. |
Discovering the Functions of Language in Online Forums (D19-55)
Copied to clipboard
| Challenge: | a vast amount of work has been dedicated to speech act categorization for characterizing discourses . lack of formalism and diversity of taxonomies make it difficult to compare different annotated datasets. |
| Approach: | They propose a semi-supervised framework for predicting the functions of Reddit comments . they propose to use the framework to analyze online forum conversations . |
| Outcome: | The proposed framework can predict functions of Reddit comments and 165K comments. |
Aligning Backchannel and Dialogue Context Representations via Contrastive LLM Fine-Tuning (2026.acl-long)
Copied to clipboard
| Challenge: | Prior work on predicting backchannel timing has focused on lexical form and prosody, but the relationship between lexico-prosodic form and meaning remains underexplored. |
| Approach: | They propose a framework for fine-tuning large language models on dialogue transcripts to derive rich contextual representations; and a joint embedding space for dialogue contexts and backchannel realizations. |
| Outcome: | The proposed framework improves context-backchannel retrieval and human perception is more sensitive to extended conversational context and embeddings align more closely with human judgments than raw WavLM features. |