Papers by Hiroaki Saito

3 papers
Let’s Put Ourselves in Sally’s Shoes: Shoes-of-Others Prefilling Improves Theory of Mind in Large Language Models (2026.findings-eacl)

Copied to clipboard

Challenge: Existing methods for Theory of Mind (ToM) are specialized for inferring beliefs from contexts involving changes in the world state.
Approach: They propose a method which makes fewer assumptions about contexts and is applicable to broader scenarios.
Outcome: The proposed method makes fewer assumptions about contexts and is applicable to broader scenarios.
Deep Reinforcement Learning with Hierarchical Action Exploration for Dialogue Generation (2024.lrec-main)

Copied to clipboard

Challenge: Existing approaches to improve dialogues with random sampling are inefficient due to the large number of eligible responses with high action values.
Approach: They propose a dual-granularity Q-function that extracts actions based on a grained hierarchy . they use offline RL and learn from multiple reward functions designed to capture emotional nuances in human interactions.
Outcome: The proposed approach outperforms baselines across automatic metrics and human evaluations.
A Personalized Dialogue Generator with Implicit User Persona Detection (2022.coling-1)

Copied to clipboard

Challenge: Existing models for personalized dialogue generation tend to be self-centered, with little care for the user in the dialogue.
Approach: They propose a personalized dialogue generator by detecting an implicit user persona and using conditional variational inference to model the user's potential persona with no external knowledge.
Outcome: The proposed model improves both automatic metrics and human evaluations by focusing on the user's persona and posterior-discriminated regularization.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations