Challenges of Using Text Classifiers for Causal Inference (D18-1)

Copied to clipboard

Challenge: a number of scientific analyses focus on low-dimensional structured data, but text classifiers can be used to produce structured variables.
Approach: They propose to use text classifiers to conduct causal analyses on simulated and Yelp data.
Outcome: The proposed method can be used on simulated and Yelp data.

Similar Papers

Causal Inference in Natural Language Processing: Estimation, Prediction, Interpretation and Beyond (2022.tacl-1)

Copied to clipboard

Challenge: causality has not had the same importance in natural language processing, says aaron e. smith . he says research on causality in NLP remains scattered across domains without unified definitions .
Approach: They propose to consolidate research on causality in NLP across academic areas . they explore potential uses of causal inference to improve robustness, fairness, interpretability .
Outcome: The proposed method is a unified overview of causal inference for the NLP community.
Causal Inference with Large Language Model: A Survey (2025.findings-naacl)

Copied to clipboard

Challenge: Existing causal inference frameworks do not match human judgment in several key areas, such as domain knowledge, logical inference, and cultural context.
Approach: They propose to apply large language models to causal inference tasks . they summarize the main causal problems and approaches and compare their results .
Outcome: The proposed methods are compared with traditional methods in healthcare, finance, and economics.
CausalNLP Tutorial: An Introduction to Causality for Natural Language Processing (2022.emnlp-tutorials)

Copied to clipboard

Challenge: Establishing causal relationships is a fundamental goal of scientific research . lack of clear definitions, notations, benchmark datasets, and challenges remains .
Approach: They introduce the fundamentals of causal discovery and causal effect estimation to the natural language processing audience and provide an overview of causal perspectives to NLP problems.
Outcome: This tutorial introduces the fundamentals of causal discovery and causal effect estimation to the natural language processing audience and provides an overview of causal perspectives to NLP problems.
A Review of Dataset and Labeling Methods for Causality Extraction (2020.coling-main)

Copied to clipboard

Challenge: Existing methods for causal relationship extraction are limited and lack of unified methods hinder progress in the field.
Approach: They propose to summarize existing methods and propose a new causal sequence label method . they propose to use multiple candidate causal label sequences according to label controversy .
Outcome: The proposed method summarises existing methods and explores their practicability and extensibility from multiple perspectives.
Text and Causal Inference: A Review of Using Text to Remove Confounding from Causal Estimates (2020.acl-main)

Copied to clipboard

Challenge: Unmeasured or latent confounders can bias causal estimates and this has motivated interest in measuring potential confounder from observed text.
Approach: They propose to use text to measure potential confounders in a way that allows for a rich measurement of multiple confounder variables.
Outcome: The proposed method is based on an individual’s entire history of social media posts or the content of a news article.
Large Language Models and Causal Inference in Collaboration: A Comprehensive Survey (2025.findings-naacl)

Copied to clipboard

Challenge: Large Language Models (LLMs) have shown great potential to enhance Natural Language Processing (NLP) models in areas such as predictive accuracy, fairness, robustness, and explainability.
Approach: They evaluate or improve generative Large Language Models from a causal perspective in areas such as reasoning capacity, fairness and safety issues, explainability, and handling multimodality.
Outcome: The proposed models can be used to perform causal relationship discovery and causal effect estimation tasks.
Pipeline for modeling causal beliefs from natural language (2023.acl-demo)

Copied to clipboard

Challenge: Existing methods to analyze language data for psychological causality are difficult to advance as they do not isolate cognitive mechanisms.
Approach: They propose a pipeline that leverages a Large Language Model to identify causal claims made in natural language documents and applies a clustering algorithm to group causal claims based on their semantic topics.
Outcome: The proposed pipeline analyzes the Covid-19 vaccine in tweets and generates a causal claim network.
DoubleLingo: Causal Estimation with Large Language Models (2024.naacl-short)

Copied to clipboard

Challenge: Existing methods for causal estimation are inadequate for noisy text data.
Approach: They propose to use LLM-based nuisance models to estimate causal effects from non-randomized data using assumptions about the underlying data distribution.
Outcome: The proposed method reduces the relative absolute error by 10.4% over existing methods on the best available dataset.
CausalEval: Towards Better Causal Reasoning in Language Models (2025.naacl-long)

Copied to clipboard

Challenge: Large language models (LLMs) have been used for a variety of tasks, including problem-solving, decision-making, and understanding of the world.
Approach: They propose a review of existing methods aimed at enhancing LMs for causal reasoning . they categorize existing methods as reasoning engines or as helpers providing knowledge or data to traditional methods .
Outcome: The proposed methods perform better than existing methods on a range of tasks.
Language Models as Causal Effect Generators (2025.emnlp-main)

Copied to clipboard

Challenge: Using sequence-driven structural causal models (SD-SCMs) we characterize how SD-SCAMs enables sampling from observational, interventional, and counterfactual distributions according to the desired causal structure.
Approach: They propose a sequence-driven structural causal model that uses language models to parameterize a structural causal system based on a user-specified DAG.
Outcome: The proposed method outperforms state-of-the-art methods and can underpin auditing of language models for (un)desirable causal effects, such as misinformation or discrimination.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations