Papers by Yuanning Feng

1 papers
Wait, We Don’t Need to “Wait”! Removing Thinking Tokens Improves Reasoning Efficiency (2025.findings-emnlp)

Copied to clipboard

Challenge: Recent advances in large reasoning models often introduce significant overthinking . this leads to verbose and redundant outputs that hinder efficiency.
Approach: They propose a plug-and-play solution that disables explicit self-reflection . it suppresses tokens such as "Wait" and "Hmm" during inference .
Outcome: The proposed approach reduces chain-of-thought trajectory length by up to 27%–51% in five R1-style model series without compromising model utility.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations