A large-scale study of the effects of word frequency and predictability in naturalistic reading (N19-1)
Copied to clipboard
| Challenge: | Recent studies have shown separable effects of word frequency and predictability on human sentence processing . other theories hold that apparent effects of frequency are underlyingly effects of predictability . |
| Approach: | They examine the generalizability of this finding to more realistic conditions of sentence processing by studying effects of frequency and predictability in three large-scale naturalistic reading corpora. |
| Outcome: | The results show that word frequency and predictability are significant in isolation but not over and above predictability, and raise doubts about the existence of such a distinction in everyday sentence comprehension. |
Similar Papers
Word Frequency Does Not Predict Grammatical Knowledge in Language Models (2020.emnlp-main)
Copied to clipboard
| Challenge: | Neural language models learn the grammatical properties of natural languages to varying degrees of accuracy. |
| Approach: | They focus on subject-verb agreement and reflexive anaphora to investigate whether there are systematic sources of variation in the language models’ accuracy. |
| Outcome: | The proposed model can learn grammatical properties from training data. |
How Furiously Can Colorless Green Ideas Sleep? Sentence Acceptability in Context (2020.tacl-1)
Copied to clipboard
| Challenge: | a recent study shows that context affects our perception of sentence acceptability, but few studies investigate how it affects language models. |
| Approach: | They compare acceptability ratings of sentences judged in isolation with a relevant context and with an irrelevant context. |
| Outcome: | The proposed model achieves state-of-the-art for unsupervised acceptability prediction. |
Measuring the Impact of (Psycho-)Linguistic and Readability Features and Their Spill Over Effects on the Prediction of Eye Movement Patterns (2022.acl-long)
Copied to clipboard
| Challenge: | Existing work to predict gaze patterns during naturalistic reading has not been conducted on general text characteristics. |
| Approach: | They propose to use two eye-tracking corpora of naturalistic reading and two language models to test their performance. |
| Outcome: | The proposed models predict eye-tracking measures during naturalistic reading and language processing. |
More than just Frequency? Demasking Unsupervised Hypernymy Prediction Methods (2021.findings-acl)
Copied to clipboard
| Challenge: | Using unsupervised methods of hypernymy prediction, we show that the predictions of three methods overlap and are highly correlated with frequency-based predictions. |
| Approach: | They compare unsupervised methods of hypernymy prediction to supervised methods . they show that the methods overlap and are highly correlated with frequency-based predictions . |
| Outcome: | The proposed methods overlap and are highly correlated with frequency-based predictions across English and German datasets. |
What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length (2025.naacl-long)
Copied to clipboard
| Challenge: | Prior work on LM and acceptability judgments treat these effects uniformly across models, making a strong assumption that models require the same degree of adjustment to control for length and unigram frequency effects. |
| Approach: | They propose a linking theory where the optimal level of adjustment is estimated from data via learned parameters for length and unigram frequency. |
| Outcome: | The proposed theory outperforms a commonly used linking theory for acceptability—SLOR—across two families of transformer LMs. |
On the Role of Context in Reading Time Prediction (2024.emnlp-main)
Copied to clipboard
| Challenge: | a new perspective on how readers integrate context during reading time prediction is presented . a recent study shows that the proportion of variance in reading times explained by context is smaller when context is represented by the orthogonalized predictor. |
| Approach: | They propose a technique where they project surprisal onto the orthogonal complement of frequency. |
| Outcome: | The proposed method shows that the proportion of variance in reading times explained by context is smaller when context is represented by the orthogonalized predictor. |
Quantifying Cognitive Factors in Lexical Decline (2021.tacl-1)
Copied to clipboard
| Challenge: | Existing studies on lexical decline suggest that cognitive and linguistic factors play a role in the survival of words and their success in the linguistic ecosystem. |
| Approach: | They propose a variety of psycholinguistic factors that are predictive of lexical decline, in which words greatly decrease in frequency over time. |
| Outcome: | The proposed factors show significant differences in the expected direction between each curated set of declining words and their matched stable words. |
Speakers enhance contextually confusable words (2020.acl-main)
Copied to clipboard
| Challenge: | Recent work has found that natural languages are shaped by pressures for efficient communication. |
| Approach: | They develop a measure of contextual confusability during word recognition based on psychoacoustic data and apply it to naturalistic speech corpora. |
| Outcome: | The proposed measure of confusability suggests that speakers alter productions to make contextually more confused words easier to understand. |
Assessing the Effect of Context in Multi-domain Acceptability Judgment (2026.findings-acl)
Copied to clipboard
| Challenge: | Existing studies evaluate sentences in isolation and do not consider how context influences LLM acceptability judgments. |
| Approach: | They examine how contextual cues affect model-generated acceptability ratings across multiple domains and several LLMs, using different forms of domain-specific contextual cueeds to situate sentences in intended usage settings. |
| Outcome: | The findings support the development of more context-aware evaluation frameworks. |
Using surprisal and fMRI to map the neural bases of broad and local contextual prediction during natural language comprehension (2021.findings-acl)
Copied to clipboard
| Challenge: | a prior work using surprisal only considered within-sentence context, using n-grams, neural language models, or syntactic structure as conditioning context. |
| Approach: | They extend the surprisal approach to use broader topical context . they identify distinct patterns of neural activation for lexical surprised and topical surpresed . |
| Outcome: | The proposed method captures effects of local and topical contexts on processing . it shows that local and broad contextual cues recruit different brain regions . |