Papers by Xiaomeng Guo

2 papers
LeTS: Learning to Think-and-Search via Process-and-Outcome Reward Hybridization (2025.emnlp-main)

Copied to clipboard

Challenge: Recent research focuses on integrating reasoning capabilities into the realm of retrieval-augmented generation (RAG) via outcome-supervised reinforcement learning (RL).
Approach: They propose a process-level reward module to mitigate the unawareness of intermediate reasoning steps in outcome-level supervision without additional annotation.
Outcome: The proposed framework can boost LLMs’ reasoning ability by integrating external knowledge sources through retrieval-augmented generation (RAG) The proposed model can mitigate the unawareness of intermediate reasoning steps in outcome-level supervision without additional annotation.
Dynamics of Cognitive Heterogeneity: Investigating Behavioral Biases in Multi-Stage Supply Chains with LLM-Based Simulation (2026.acl-long)

Copied to clipboard

Challenge: Recent advances in Large Language Models (LLMs) allow repeatable experiments in which individual characteristics can be precisely defined.
Approach: They propose a scalable experimental paradigm using Large Language Models to simulate multi-stage supply chain dynamics.
Outcome: The proposed model systematically replicates and validates the results of a behavioral simulation on agents in multi-stage supply chain dynamics.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations