The time scale of redundancy between prosody and linguistic context (2025.acl-long)
Copied to clipboard
Tamar I Regev, Chiebuka Ohams, Shaylee Xie, Lukas Wolf, Evelina Fedorenko, Alex Warstadt, Ethan Wilcox, Tiago Pimentel
| Challenge: | Prior work has shown that the information carried by prosodic features is substantially redundant with that carried by the surrounding words. |
| Approach: | They examine the time scale of this relationship, studying how it varies with the length of past and future contexts. |
| Outcome: | The results show that prosody features show some redundancy with future words, but only with a short scale of 1-2 words, consistent with reports of incremental short-term planning in language production. |
Similar Papers
Quantifying the redundancy between prosody and text (2023.emnlp-main)
Copied to clipboard
Lukas Wolf, Tiago Pimentel, Evelina Fedorenko, Ryan Cotterell, Alex Warstadt, Ethan Wilcox, Tamar Regev
| Challenge: | Existing studies suggest partial redundancy between prosody and linguistic information. |
| Approach: | They use large language models to estimate how much information is redundant between prosody and the words themselves. |
| Outcome: | The proposed model can predict prosodic features across prosodic features, including intensity, duration, pauses, and pitch contours. |
What Do Prosody and Text Convey? Characterizing How Meaningful Information is Distributed Across Multiple Channels (2026.acl-long)
Copied to clipboard
| Challenge: | Prosody—the melody of speech—conveys critical information often not captured by the words or text of a message. |
| Approach: | They propose an information-theoretic approach to quantify how much is conveyed by prosody that is not recoverable from text alone. |
| Outcome: | The proposed framework can quantify how much is conveyed by prosody that is not recoverable from text alone and crucially, what prosody conveys. |
Using Information Theory to Characterize Prosodic Typology: The Case of Tone, Pitch-Accent and Stress-Accent (2025.acl-long)
Copied to clipboard
| Challenge: | lexical identity and prosody are well-studied parameters of linguistic variation, but they are difficult to predict in tonal languages. |
| Approach: | They propose to characterize the relationship between lexical identity and prosody using information theory to estimate mutual information between the text and pitch curves. |
| Outcome: | The proposed hypothesis supports perspectives that view linguistic typology as gradient, rather than categorical. |
Prosody: Models, Methods, and Applications (2021.acl-tutorials)
Copied to clipboard
| Challenge: | This tutorial will overview the computational modeling of prosody. |
| Approach: | This tutorial will overview the computational modeling of prosody . it will discuss the latest advances in prosody and diverse applications . |
| Outcome: | This tutorial will overview the computational modeling of prosody. |
The Prosody of Emojis (2026.acl-long)
Copied to clipboard
| Challenge: | emojis are useful in spoken communication because they add affective and pragmatic nuance to textual cues. |
| Approach: | They analyze human speech data to find prosodic features that are important in spoken communication. |
| Outcome: | The proposed model shows that speakers adapt prosody based on emoji cues, and that listeners can recover intended meanings significantly above chance. |
Giving Attention to the Unexpected: Using Prosody Innovations in Disfluency Detection (N19-1)
Copied to clipboard
| Challenge: | Disfluencies in spontaneous speech are associated with prosodic disruptions. |
| Approach: | They propose a method to extract acoustic-prosodic cues from word transcripts . they explore early and late fusion techniques for integrating text and prosody . |
| Outcome: | The proposed approach shows gains over a high-accuracy text-only model. |
Does Context Matter? A Prosodic Comparison of English and Spanish in Monolingual and Multilingual Discourse Settings (2025.emnlp-main)
Copied to clipboard
| Challenge: | a large number of studies on prosody in languages have focused on monolingual discourse contexts . a recent study focused on the prosodic features of monolingual speech in multilingual contexts. |
| Approach: | They compare prosody of monolingual English and Spanish in monolingual and multilingual settings . they find that monolingual speech produced in a monolingual context is prosodically different from that produced in multilingual context . |
| Outcome: | The proposed study is the first to incorporate multilingual discourse contexts into the study of native-level monolingual prosody. |
On the Role of Context in Reading Time Prediction (2024.emnlp-main)
Copied to clipboard
| Challenge: | a new perspective on how readers integrate context during reading time prediction is presented . a recent study shows that the proportion of variance in reading times explained by context is smaller when context is represented by the orthogonalized predictor. |
| Approach: | They propose a technique where they project surprisal onto the orthogonal complement of frequency. |
| Outcome: | The proposed method shows that the proportion of variance in reading times explained by context is smaller when context is represented by the orthogonalized predictor. |
The Role of Prosody in Spoken Question Answering (2025.findings-naacl)
Copied to clipboard
| Challenge: | lexical information is not available in most models, but prosody is important in understanding spoken language. |
| Approach: | They investigate the role of prosody in the process of answering a spoken question by isolating prosodic and lexical information from a natural speech dataset. |
| Outcome: | The proposed models can perform reasonably well on the SLUE-SQA-5 dataset, but when lexical information is available, models tend to predominantly rely on it. |
A surprisal–duration trade-off across and within the world’s languages (2021.emnlp-main)
Copied to clipboard
| Challenge: | Throughout human evolution, countless languages have evolved, each with unique features. |
| Approach: | They analysed a corpus of 600 languages to find strong evidence for a surprisal–duration trade-off between languages and languages. |
| Outcome: | The proposed model shows that phones are produced faster in languages where they are less surprising and vice versa. |