Papers by Zhenyun Yin
Deliberative Searcher: Improving LLM Reliability via Reinforcement Learning with Constraints (2026.acl-long)
Copied to clipboard
| Challenge: | Large language models with search capabilities often exhibit miscalibrated confidence, causing incorrect answers with high certainty. |
| Approach: | They propose a reasoning-primary framework that integrates search operations into chain-of-thought generation while maintaining explicit confidence calibration. |
| Outcome: | The proposed framework improves accuracy and reliability of large language models with search capabilities. |