SocialRL Enables Strategic Negotiation in 4B Parameter Models
August 17, 2026
SocialRL is a training recipe that improves social reasoning in small language models, enabling them to act as strategic agents rather than passive assistants. A 4B model trained with this method matched or exceeded frontier models on held-out scenarios across six domains including marketplaces and job interviews.
HOW THIS AFFECTS YOU
●
builderYou can deploy smaller, more cost-effective models for complex agentic tasks like negotiation and haggling.
●
researcherThe recipe shows that social reasoning can be effectively distilled into small-scale architectures through specialized RL.