Papers by Anna Marshalova

1 papers
Call, Reward, Repeat: Advancing Dialog State Tracking with GRPO and Function Calling (2026.eacl-srw)

Copied to clipboard

Challenge: Recent advances in Large Language Models (LLMs) have notably enhanced task-oriented dialogue systems, particularly in Dialogue State Tracking (DST).
Approach: They propose a group-relative policy optimization method that guides LLMs toward improved DST accuracy even under low-resource conditions.
Outcome: The proposed method improves on established DST benchmarks while using significantly reduced out-of-domain training data.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations