Papers by Seungwon Lim

5 papers
Persona Dynamics: Unveiling the Impact of Persona Traits on Agents in Text-Based Games (2025.acl-long)

Copied to clipboard

Challenge: Text-based interactive environments have long presented formidable challenges for AI.
Approach: They propose a method for projecting human personality traits onto agents to guide their behavior and integrate them into their policy-learning pipelines.
Outcome: The proposed method induces personality in a text-based game agent by integrating personality profiles directly into the agent's policy-learning pipeline.
Can visual language models resolve textual ambiguity with visual cues? Let visual puns tell you! (2024.emnlp-main)

Copied to clipboard

Challenge: Existing models lack this active understanding capacity, limiting their applicability in real-world scenarios.
Approach: They propose a benchmark to assess the impact of multimodal inputs on lexical ambiguities.
Outcome: The proposed benchmark assesses the impact of multimodal inputs on lexical ambiguities.
PAC-BENCH: Evaluating Multi-Agent Collaboration under Privacy Constraints (2026.findings-acl)

Copied to clipboard

Challenge: Recent research explores multi-agent systems where agents collaborate toward shared goals to handle complex tasks.
Approach: They propose a benchmark for systematic evaluation of multi-agent collaboration under privacy constraints.
Outcome: The proposed benchmark shows that privacy constraints degrade collaboration performance and make outcomes depend more on the initiating agent than the partner.
Do LLMs Have Distinct and Consistent Personality? TRAIT: Personality Testset designed for LLMs with Psychometrics (2025.findings-naacl)

Copied to clipboard

Challenge: Recent advances in Large Language Models (LLMs) have led to their adaptation as conversational agents.
Approach: They propose a new benchmark that uses 8K multi-choice questions to assess the personality of Large Language Models.
Outcome: The proposed personality test outperforms existing personality tests for LLMs in reliability and validity.
VisEscape: A Benchmark for Evaluating Exploration-driven Decision-making in Virtual Escape Rooms (2025.emnlp-main)

Copied to clipboard

Challenge: Existing studies on embodied agents have addressed the importance of exploration in environments where tasks and solutions are not predefined.
Approach: They propose a virtual escape room that evaluates AI models in a dynamic environment . they propose to integrate memory management and reasoning into the simulation .
Outcome: The proposed model improves in dynamic and exploration-driven environments by integrating memory management and reasoning.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations