Papers by Zimo Qi
Discourse Heuristics For Paradoxically Moral Self-Correction (2025.findings-emnlp)
Copied to clipboard
| Challenge: | moral self-correction is a promising approach for aligning output of Large Language Models with human moral values . authors show that moral self correction relies on discourse constructions that reflect heuristic shortcuts . |
| Approach: | a new method is proposed to strengthen moral self-correction using heuristics extracted from curated datasets. |
| Outcome: | a new method to strengthen moral self-correction is proposed . the proposed method is based on heuristics extracted from curated datasets. |
Diagnosing Moral Reasoning Acquisition in Language Models: Pragmatics and Generalization (2025.findings-emnlp)
Copied to clipboard
| Challenge: | Prior research has shown that LLMs fail to perform satisfactorily on moral cognizance tasks . |
| Approach: | They propose to use curated datasets to improve LLMs' moral cognizance . they find pragmatic dilemma constrains generalization ability of current learning paradigms . |
| Outcome: | The proposed learning paradigms fail to perform on moral cognizance tasks, the authors show . they show that the pragmatic dilemma is the primary bottleneck for moral reasoning acquisition . |