Papers by Thin Nguyen

2 papers
Causal Direct Preference Optimization for Language Model Alignment (2026.findings-eacl)

Copied to clipboard

Challenge: Empirical evaluations show that CDPO surpasses DPO-based baselines by achieving unbiased fine-tuning through causal reasoning.
Approach: They propose a framework that incorporates causal inference principles to mitigate the influence of confounders and sharpen the signal of genuine human preferences.
Outcome: The proposed framework preserves the tractability of direct optimization while enhancing robustness to spurious correlations and annotation biases.
Causal Activation Steering via Sparse Mediation (2026.findings-eacl)

Copied to clipboard

Challenge: a sparse mediation steering approach to control language-model behavior is feasible, says a new study . existing methods that learn dense steering vectors modify thousands of activation dimensions simultaneously .
Approach: They propose a sparse mediation steering approach that learns targeted behavioral interventions via regularized training.
Outcome: The proposed method achieves 97-100% of dense baseline effectiveness across four tasks while using only 10-30% of activation dimensions.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations