Papers by Michael Bowling

1 papers
KETCHUP: K-Step Return Estimation for Sequential Knowledge Distillation (2026.findings-eacl)

Copied to clipboard

Challenge: Empirical evaluation shows that our approach yields superior performance in both standard task metrics and large language model (LLM)-based evaluation.
Approach: They propose a K-step return estimation method for reinforcement learning (RL)-based knowledge distillation in text generation tasks using the Bellman Optimality Equation.
Outcome: The proposed method performs better on standard task metrics and large language model evaluations on three text generation tasks.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations