Papers by Yerong Wu
StratMem-Bench: Evaluating Strategic Memory Use in Virtual Character Conversation Beyond Factual Recall (2026.acl-long)
Copied to clipboard
| Challenge: | Current benchmarks for memory utilization ignore this nuance, treating memory as a static repository of facts rather than a dynamic resource to be strategically deployed in character-centric dialogues. |
| Approach: | They propose a benchmark to evaluate strategic memory use in character-centric dialogues . they use a dataset of 657 instances where virtual characters must navigate heterogeneous memory pools . |
| Outcome: | The proposed benchmarks show that all models perform well at distinguishing between required and irrelevant memories, but struggle once supportive memories are introduced into the decision process. |