Papers by Yexin Wu

2 papers
Are the Values of LLMs Structurally Aligned with Humans? A Causal Perspective (2025.findings-acl)

Copied to clipboard

Challenge: Current approaches to value alignment focus on a few core values, such as helpfulness, harmlessness, and honesty.
Approach: They propose to use latent causal value graphs to guide two lightweight value-steering methods . role-based prompting and sparse autoencoder (SAE) steering are also used .
Outcome: Experiments on Gemma-2B-IT and Llama3-8B- IT show that the proposed methods are effective and controllable.
Mitigating Misleading Chain-of-Thought Reasoning with Selective Filtering (2024.lrec-main)

Copied to clipboard

Challenge: Large language models have demonstrated remarkable capabilities by leveraging chain-of-thought reasoning techniques to solve complex questions.
Approach: They propose a method that assesses the entailment relationship between the question and the candidate reasoning chain and uses it to predict the answer.
Outcome: The proposed approach improves the fine-tuned T5 baseline over the ScienceQA, ECQA, and LastLetter tasks.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations