DIPLOMAT: Dialogue-Span-Aware DPO for Workplace Negotiation
September 22, 2026
DIPLOMAT uses Dialogue-Span-Aware Direct Preference Optimization to train models for polite and persuasive workplace negotiations. The approach is supported by PROWESS, a new multi-agent generated dataset containing multi-turn dialogues with labeled negotiation and politeness strategies.
HOW THIS AFFECTS YOU
●
builderThis provides a new method for fine-tuning agents to handle complex, multi-objective social interactions.
●
researcherThe PROWESS dataset offers a new benchmark for evaluating strategy-aligned dialogue agents.