Papers by Lexiang Tang

1 papers
LEASH: Adaptive Length Penalty and Reward Shaping for Efficient Large Reasoning Model (2026.acl-long)

Copied to clipboard

Challenge: Existing approaches to long reasoning traces are hard to tune and fail to adapt to evolving LLMs.
Approach: They propose a reinforcement learning framework that optimizes the length of reasoning traces by a Lagrangian primal–dual method.
Outcome: The proposed framework reduces the average reasoning length by 60% across diverse tasks while maintaining competitive performance.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations