ASTRA: A Negotiation Agent with Adaptive and Strategic Reasoning via Tool-integrated Action for Dynamic Offer Optimization (2025.emnlp-main)
Copied to clipboard
| Challenge: | Existing agents struggle due to bounded rationality in human data, low adaptability to counterpart behavior, and limited strategic reasoning. |
| Approach: | They propose a framework for turn-level offer optimization based on two core principles: opponent modeling and Tit-for-Tat reciprocity. |
| Outcome: | The proposed framework outperforms baselines across diverse partner agents and validates through human evaluation. |
Similar Papers
A Dual-Mind Framework for Strategic and Expressive Negotiation Agent (2025.acl-long)
Copied to clipboard
| Challenge: | Existing approaches to negotiation dialogue focus on only one aspect, ignoring the synergistic effect of their combined synergies. |
| Approach: | They propose a dual-mind negotiation agent framework that integrates an intuitive and a deliberative module for slow, expression optimization. |
| Outcome: | The proposed framework achieves state-of-the-art on negotiation datasets showing that it improves negotiation ability. |
INA: An Integrative Approach for Enhancing Negotiation Strategies with Reward-Based Dialogue Agent (2023.findings-emnlp)
Copied to clipboard
| Challenge: | a novel negotiation agent is designed for the online marketplace . a dialogue agent can negotiate on price and other factors . |
| Approach: | They propose a novel negotiation agent that is integrative in nature and can negotiate on price and other factors. |
| Outcome: | The proposed agent is integrative in nature and can negotiate on price and other factors. |
PRISMA: Preference-Reinforced Self-Training Approach for Interpretable Emotionally Intelligent Negotiation Dialogues (2026.acl-long)
Copied to clipboard
| Challenge: | Emotion plays a pivotal role in shaping negotiation outcomes, influencing trust, cooperation, and long-term relationships. |
| Approach: | They propose an Emotion-aware Negotiation Strategy-informed Chain-of-Thought reasoning mechanism which mimics human negotiation by perceiving, understanding, using, and managing emotions. |
| Outcome: | The proposed system generates interpretable emotions and improves negotiation effectiveness on job interviews and resource allocation datasets. |
ASTRA: An Automated Framework for Strategy Discovery, Retrieval, and Evolution for Jailbreaking LLMs (2026.acl-long)
Copied to clipboard
| Challenge: | Existing methods lack the capability for continuous learning and self-evolution from interactions, limiting the diversity and adaptability of attack strategies. |
| Approach: | They propose an automated framework capable of discovering, retrieving, and evolving attack strategies. |
| Outcome: | The proposed framework outperforms existing baselines in a black-box setting. |
Decoupling Strategy and Generation in Negotiation Dialogues (D18-1)
Copied to clipboard
| Challenge: | Recent work on negotiation trains neural models, but their end-to-end nature makes it hard to control their strategy. |
| Approach: | They propose a modular approach that decouples strategy and generation by coarse dialogue acts . they test their approach on a recently proposed DEALORNODEAL game . |
| Outcome: | The proposed approach can decouple strategy and generation without degeneracy. |
EmoMAS: Emotion-Aware Multi-Agent System for High-Stakes Edge-Deployable Negotiation with Bayesian Orchestration (2026.acl-long)
Copied to clipboard
| Challenge: | Large language models (LLMs) are increasingly used for automated negotiation, but their cloud-centric paradigm exposes sensitive negotiations to privacy and security risks. |
| Approach: | They propose a Bayesian multi-agent framework that transforms emotional decision-making from reactive to strategic. |
| Outcome: | EmoMAS leverages a Bayesian orchestrator to coordinate three specialized agents: game-theoretic, reinforcement learning, and psychological coherence models. |
Improving Dialog Systems for Negotiation with Personality Modeling (2021.acl-long)
Copied to clipboard
| Challenge: | In this paper, we introduce a framework for generating strategic dialog inspired by the idea of incorporating a theory of mind (ToM) into machines. |
| Approach: | They propose a probabilistic formulation to encapsulate the opponent's personality type during both learning and inference. |
| Outcome: | The proposed model achieves 20% higher dialog agreement rate compared to baselines on a mixed population of opponents. |
PrefIx: Understand and Adapt to User Preference in Human-Agent Interaction (2026.findings-acl)
Copied to clipboard
| Challenge: | Current benchmarks evaluate task accuracy but overlook how agents interact . Preference-aware agents show 7.6% average UX improvement and 18.5% gain in preference alignment. |
| Approach: | They propose a configurable environment that evaluates both what agents accomplish and how they interact. |
| Outcome: | The proposed model improves performance and improves user experience by 7.6% and 18.5% respectively. |
Beyond Static Testbeds: An Interaction-Centric Agent Simulation Platform for Dynamic Recommender Systems (2025.emnlp-main)
Copied to clipboard
| Challenge: | Existing platforms lack a mechanism for user actions to dynamically reshape the environment. |
| Approach: | They propose a novel agent-based simulation platform for recommender systems with a robust interaction mechanism. |
| Outcome: | The proposed platform improves the credibility of the simulation and replicates the Matthew Effect and Brand Loyalty. |
Beyond Task-Oriented and Chitchat Dialogues: Proactive and Transition-Aware Conversational Agents (2025.emnlp-main)
Copied to clipboard
| Challenge: | Current efforts to bridge the two modes of interaction are reactive, focusing on responding to user inputs rather than coordinating dialogue flows. |
| Approach: | They propose a dataset designed for transition-aware dialogue modeling that incorporates structurally diverse and integrated mode flows. |
| Outcome: | The proposed dataset outperforms baseline models in intent detection and mode transition handling. |