A Cognitive Regularizer for Language Modeling (2021.acl-long)

Copied to clipboard

Challenge: a uniform information density hypothesis is used to explain certain linguistic phenomena . a regularizer that encodes the UID hypothesis can be used for language training .
Approach: They propose to augment the canonical MLE objective with a regularizer that encodes UID . they find that regularization consistently improves perplexity in language models .
Outcome: The proposed hypothesis can be operationalized as an inductive bias for language modeling.

Similar Papers

Revisiting the Uniform Information Density Hypothesis (2021.emnlp-main)

Copied to clipboard

Challenge: The uniform information density hypothesis posits a preference among language users for utterances structured such that information is distributed uniformly across a signal.
Approach: They propose to test the hypothesis by using reading time and acceptability data to examine the effect of surprisal on language comprehension and acceptabilities.
Outcome: The proposed hypothesis makes predictions about language comprehension and linguistic acceptability .
Revisiting the Uniform Information Density Hypothesis in LLM Reasoning (2026.findings-acl)

Copied to clipboard

Challenge: a recent study has highlighted the fragility of Chain-of-Thought reasoning . a hypothesis suggests that effective communication is achieved by maintaining a stable flow of information.
Approach: They propose a framework to quantify uniformity of information flow at local and global levels . they propose entropy-based stepwise density metric to quantify this phenomenon .
Outcome: The proposed framework outperforms alternative signals as predictors of reasoning quality.
Revisiting Entropy Rate Constancy in Text (2023.findings-emnlp)

Copied to clipboard

Challenge: Existing evidence supports the uniform information density hypothesis . however, we re-evaluate the hypothesis with neural language models .
Approach: They propose to use n-gram language models to argue that English documents exhibit entropy rate constancy . they re-evaluate the claims of Genzel and Charniak with neural language models .
Outcome: The proposed hypothesis fails to support the proposed hypothesis with language models.
Surprise! Uniform Information Density Isn’t the Whole Story: Predicting Surprisal Contours in Long-form Discourse (2024.emnlp-main)

Copied to clipboard

Challenge: Uniform Information Density (UID) hypothesis posits that speakers tend to distribute information evenly across linguistic units to achieve efficient communication.
Approach: They propose a functional pressure that speakers modulate information rate based on location within a hierarchically-structured model of discourse.
Outcome: The proposed hypothesis posits that speakers tend to distribute information evenly across linguistic units to achieve efficient communication.
Is Information Density Uniform when Utterances are Grounded on Perception and Discourse? (2026.eacl-long)

Copied to clipboard

Challenge: Existing studies on the distribution of information in visually grounded contexts have focused on text-only inputs.
Approach: They propose to use multilingual vision-and-language models to estimate surprisal . they find grounding on perception increases uniformity across typologically diverse languages .
Outcome: The proposed hypothesis is tested in visual-language models over 30 languages and 13 storytelling languages . the results show grounding on perception increases uniformity across languages compared to text-only settings .
GPT-who: An Information Density-based Machine-Generated Text Detector (2024.findings-naacl)

Copied to clipboard

Challenge: Large Language Models (LLMs) generate misinformation, memorized content, plagiarized content, toxic speech, and hallucinated content.
Approach: They propose a statistical detector that uses UID to model the unique statistical signature of each LLM and human author for accurate detection.
Outcome: The proposed method outperforms state-of-the-art detectors by over 20% across domains.
How do decoding algorithms distribute information in dialogue responses? (2023.findings-eacl)

Copied to clipboard

Challenge: Using different decoding algorithms, we find that human dialogue generation is beneficial for adherence to the Uniform Information Density principle.
Approach: They investigate whether decoding algorithms implicitly follow the Uniform Information Density principle by distributing information evenly in utterances.
Outcome: The proposed method encourages non-uniform responses, but under low/high surprisal conditions, resulting in poor quality responses.
A Cross-Linguistic Pressure for Uniform Information Density in Word Order (2023.tacl-1)

Copied to clipboard

Challenge: a recent study has compared real and counterfactual word orders, but one functional pressure has been overlooked . a study of 10 typologically diverse languages shows that real word orders have greater uniformity than reverse word orders .
Approach: They propose to test whether a pressure for UID may have influenced word order patterns cross-linguistically.
Outcome: The proposed model shows that real orders have greater uniformity than reverse orders among SVO languages.
Uniform Information Density and Syntactic Reduction: Revisiting *that*-Mentioning in English Complement Clauses (2025.emnlp-main)

Copied to clipboard

Challenge: Uniform Information Density (UID) hypothesis suggests that speakers exploit this variability to maintain a consistent rate of information transmission during language production.
Approach: They propose that speakers exploit this variability to maintain a consistent rate of information transmission during language production.
Outcome: The proposed hypothesis replicates the established relationship between information density and *that*-mentioning .
The Harmonic Structure of Information Contours (2025.acl-long)

Copied to clipboard

Challenge: Language typically does not maintain a uniform information rate, but it fluctuates around a global average . a new study suggests periodicity may be a factor in information rate oscillations .
Approach: They propose a hypothesis that language does not maintain a uniform information rate . they apply harmonic regression and introduce a new extension to detect periodicity .
Outcome: The proposed method reveals that language oscillates at periodic intervals across frequencies . it also offers a framework for uncovering structural pressures at various levels of linguistic granularity.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations