Challenge: Using data from news datasets, we examine readers' veridicality judgments to news events at sentence level.
Approach: They collect and study Chinese readers’ veridicality judgments to news events . goal is to observe pragmatic behaviors of linguistic features under context .
Outcome: The aim is to observe the pragmatic behaviors of linguistic features under context which affects readers in making veridicality judgments.

Similar Papers

Revisiting Classical Chinese Event Extraction with Ancient Literature Information (2025.acl-long)

Copied to clipboard

Challenge: Existing studies on classical Chinese event extraction focus on grafting the complex modeling from English or modern Chinese works, neglecting the unique characteristic of this language.
Approach: They propose a Literary Vision-Language Model (VLM) for classical Chinese event extraction . they integrate annotations, historical background and character glyphs to capture the inner- and outer-context information from the sequence.
Outcome: The proposed model can capture the inner- and outer-context information at nearly zero cost.
Identifying and Understanding User Reactions to Deceptive and Trusted Social News Sources (P18-2)

Copied to clipboard

Challenge: a new study examines how users react to news sources with different levels of credibility . a recent study found that 59% of bitly-URLs on Twitter are shared without ever being read .
Approach: They develop a model to classify user reactions into one of nine types . they also measure the speed and type of reaction for trusted and deceptive news sources .
Outcome: The proposed model classifies user reactions into one of nine types, such as answer, elaboration, and question, etc.
Pragmatics in the Era of Large Language Models: A Survey on Datasets, Evaluation, Opportunities and Challenges (2025.acl-long)

Copied to clipboard

Challenge: linguistics studies how context influences meaning of language and how people use it to convey implied meanings, emotions, and intentions.
Approach: They analyze task designs, data collection methods, evaluation approaches and their relevance to real-world applications.
Outcome: The findings highlight emerging trends, challenges, and gaps in existing benchmarks . the findings will contribute to more nuanced and context-aware NLP models .
Doc2EDAG: An End-to-End Document-level Framework for Chinese Financial Event Extraction (D19-1)

Copied to clipboard

Challenge: Existing event extraction methods are limited to extract event arguments within the sentence scope.
Approach: They propose a model which generates an entity-based directed acyclic graph to fulfill document-level EE effectively.
Outcome: The proposed model can generate entity-based directed acyclic graph to fulfill document-level EE effectively.
Determining Event Durations: Models and Error Analysis (N18-2)

Copied to clipboard

Challenge: a crucial piece of information regarding events is their duration, a rarely mentioned attribute . core tasks such as temporal understanding and reasoning would benefit from knowing the expected duration of events.
Approach: They introduce aspectual features that capture deeper linguistic information . they also experiment with neural networks to predict event durations .
Outcome: The proposed models capture deeper linguistic information than previous work and provide useful clues.
Forecasting Future International Events: A Reliable Dataset for Text-Based Event Modeling (2024.findings-emnlp)

Copied to clipboard

Challenge: Existing approaches for text-based event prediction are limited in quality due to dynamic nature of international relations and conflicting economic dynamics.
Approach: They propose a novel dataset that leverages the advanced reasoning capabilities of large-language models to address these limitations.
Outcome: The proposed dataset features high-quality scoring labels generated through advanced prompt modeling and rigorously validated by domain experts in political science.
He Thinks He Knows Better than the Doctors: BERT for Event Factuality Fails on Pragmatics (2021.tacl-1)

Copied to clipboard

Challenge: Existing models for factuality prediction are lacking for English . Traditionally, event factualism is triggered by fixed properties of lexical items .
Approach: They propose a model that exploits common surface patterns that correlate with factuality labels.
Outcome: The proposed model achieves the best performance on four factuality datasets.
Ask Again, Then Fail: Large Language Models’ Vacillations in Judgment (2024.acl-long)

Copied to clipboard

Challenge: Existing large language models often waver in their judgments when faced with follow-up questions . this is a challenge for generating reliable responses and building user trust .
Approach: They propose a Follow-up Questioning Mechanism and two metrics to quantify this inconsistency . they also develop a framework that teaches large language models to maintain original correct judgments .
Outcome: The proposed framework improves the general capabilities of large language models by allowing them to maintain original correct judgments.
A Psycholinguistic Evaluation of Language Models’ Sensitivity to Argument Roles (2024.findings-emnlp)

Copied to clipboard

Challenge: a systematic evaluation of large language models' sensitivity to argument roles is presented . a recent study shows that argument roles have a delayed impact on verb prediction in human sentence processing.
Approach: They propose to replicate psycholinguistic studies on human argument role processing . they find that language models are able to distinguish verbs that appear in plausible and implausible contexts .
Outcome: The proposed models are able to distinguish verbs that appear in plausible and implausible contexts, but none captures the same selective patterns that human comprehenders exhibit during real-time verb prediction.
Not quite Sherlock Holmes: Language model predictions do not reliably differentiate impossible from improbable events (2025.findings-acl)

Copied to clipboard

Challenge: Existing work has shown that language models can select the most likely or plausible of a set of possible events, but they are far from robust.
Approach: They focus on whether language models can select the most likely or plausible of a set of possibilities and compare them to a broader behavior that humans exhibit largely unconsciously.
Outcome: The proposed models perform worse than expected under certain conditions, compared with Llama 3, Gemma 2, and Mistral NeMo, and they are significantly more sensible than leaves.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations