Papers by Jinbiao Wei
Rethinking Reasoning-Intensive Retrieval: Evaluating and Advancing Retrievers in Agentic Search Systems (2026.acl-long)
Copied to clipboard
| Challenge: | Existing evaluation benchmarks for retrievers are narrow and evaluate them in isolation . existing evaluation benchmarking frameworks focus on evaluating retrievers in isolation, obscuring their value in real-world applications. |
| Approach: | They propose an evaluation framework that evaluates retrievers in agentic search systems . they provide expert-annotated reasoning aspects, positive documents, a reference response and evaluation rubrics . |
| Outcome: | The proposed framework assesses retrievers in agentic search systems. |
Anchor: Branch-Point Data Generation for GUI Agents (2026.acl-long)
Copied to clipboard
| Challenge: | Existing GUI agents for real desktop environments require large amounts of high-quality interaction data, but collecting human demonstrations is expensive. |
| Approach: | They propose a framework that bootstraps scalable desktop supervision from seed demonstrations. |
| Outcome: | Experiments on standard desktop benchmarks show that the framework improves on zero-shot agents and representative synthesis baselines. |