Papers by Renxuan Tan
Beyond Compromise: Pareto-Lenient Consensus for Efficient Multi-Preference LLM Alignment (2026.findings-acl)
Copied to clipboard
| Challenge: | Recent approaches to align LLMs with diverse human values are based on static linear scalarization or rigid gradient projection . however, these approaches often sacrifice potential global Pareto improvements to avoid transient local trade-offs. |
| Approach: | They propose a game-theoretic framework that reimagines alignment as a dynamic negotiation process. |
| Outcome: | The proposed framework breaks the deadlock between static linear scalarization and rigid gradient projection . it allows the model to escape local degradation and explore the distal Pareto-optimal frontier . |