On Positivity Bias in Negative Reviews (2021.acl-short)

Copied to clipboard

Challenge: Existing studies have shown positive words are more frequently used in negative reviews . however, it remains unclear whether the Pollyanna hypothesis holds in negative review .
Approach: They validate the Pollyanna hypothesis that positive words occur more frequently than negative words in human expressions . they use a variety of review datasets to examine the use of positive and negative words .
Outcome: The results confirm the pollyanna hypothesis that positive words occur more frequently than negative words in human expressions.

Similar Papers

It Is Not Only the Negative that Deserves Attention! Understanding, Generation & Evaluation of (Positive) Moderation (2025.naacl-long)

Copied to clipboard

Challenge: Moderation is essential for maintaining and improving the quality of online discussions.
Approach: They annotate a dataset on 13 modes of discussion and use it to generate positive moderation.
Outcome: The proposed model shows that professional moderation generates higher ratings than professional moderated moderation, but prefers professional moderate in pairwise comparison.
Not All Reviews Are Equal: Towards Addressing Reviewer Biases for Opinion Summarization (P19-2)

Copied to clipboard

Challenge: Existing research focuses on mining for opinions from review texts and ignores reviewers.
Approach: They propose to model reviewer biases from review texts and learn a bias-aware opinion representation.
Outcome: The proposed method includes balanced opinions from reviewers with different biases and preferences.
Mind Your Bias: A Critical Review of Bias Detection Methods for Contextual Language Models (2022.findings-emnlp)

Copied to clipboard

Challenge: Existing methods for detection of biases in contextual language models are inconsistent and inconclusive.
Approach: They propose to use word embedding association test to detect biases in contextual language models to compare them with other methods.
Outcome: The proposed methods are inconsistent and inconclusive for language models with word embeddings.
Say What You Mean! Large Language Models Speak Too Positively about Negative Commonsense Knowledge (2023.acl-long)

Copied to clipboard

Challenge: Large language models (LLMs) have been studied for their ability to store and utilize positive knowledge.
Approach: They propose to use a constrained keywords-to-sentence generation task and a Boolean question answering task to probe large language models on negative commonsense knowledge.
Outcome: The proposed tasks show that LLMs fail to generate valid sentences grounded in negative commonsense knowledge, yet they can correctly answer yes-or-no questions.
Revisiting subword tokenization: A case study on affixal negation in large language models (2024.naacl-long)

Copied to clipboard

Challenge: Negation is central to language understanding but is not properly captured by modern NLP methods.
Approach: They propose to use subword tokenization methods to detect negation in large language models . they find that models can reliably recognize negation, despite mismatches in tokenization accuracy .
Outcome: The proposed models can detect negation in English using subword tokenization methods despite some mismatches in tokenization accuracy and negation detection performance.
A Study of Nationality Bias in Names and Perplexity using Off-the-Shelf Affect-related Tweet Classifiers (2024.emnlp-main)

Copied to clipboard

Challenge: Recent research shows that named entities influence PLMs in many applications.
Approach: They propose a method to quantify biases associated with named entities from various countries using Twitter data instead of templates or specific datasets.
Outcome: The proposed method shows positive biases related to the language spoken in a country across all classifiers.
Reducing Sentiment Bias in Language Models via Counterfactual Evaluation (2020.findings-emnlp)

Copied to clipboard

Challenge: Language modeling has advanced rapidly due to efficient model architectures and the availability of large text corpora.
Approach: They propose to embed and regularize sentiment prediction-derived regularizations on the language model’s latent representations to reduce bias in the sentiment of generated text.
Outcome: The proposed methods reduce bias in the sentiment of generated text by adopting individual and group fairness metrics from the fair machine learning literature.
“Was it “stated” or was it “claimed”?: How linguistic bias affects generative language models (2021.emnlp-main)

Copied to clipboard

Challenge: Several studies have identified such linguistic classes of words that occur frequently in natural language text and are bias-inducing by virtue of their framing effects.
Approach: They propose to use linguistic cues to induce subtle biases through implied sentiment and presupposed facts to influence the distribution of the generated text.
Outcome: The proposed models are sensitive to these framing effects, but show that they lead to measurable style and topic differences in the generated text, leading to language that is, on average, more polarised and more skewed towards controversial entities and events.
An Analysis of Negation in Natural Language Understanding Corpora (2022.acl-short)

Copied to clipboard

Challenge: Using annotator-generated examples, one can evaluate systems with synthetic language that is not representative of language in the wild.
Approach: They analyze negation in eight popular corpora spanning six natural language understanding tasks.
Outcome: The proposed corpora have few negations compared to general-purpose English and are often unimportant . state-of-the-art transformers obtain significantly worse results with instances that contain negation, especially if the negations are important.
Opinions Are Not Always Positive: Debiasing Opinion Summarization with Model-Specific and Model-Agnostic Methods (2024.lrec-main)

Copied to clipboard

Challenge: Existing opinion summarization frameworks are reluctant to generate negative summaries given input of negative opinions.
Approach: They propose to disentangle input into sentiment-relevant and sentiment-irrelevant components through adversarial loss.
Outcome: The proposed approaches reduce sentiment bias in the existing opinion summarization dataset . the proposed approaches generate better summaries with a more balanced emotional polarity distribution .

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations