CAVEAT Benchmark Reveals Agent Vulnerability to Misaligned Marketplace Incentives
September 24, 2026
The CAVEAT benchmark tests computer-use agents against environments designed to steer them away from user objectives. In testing, agent success in selecting user-optimal products dropped from 78.6% to 17.3% when encountering marketplace steering mechanisms.
HOW THIS AFFECTS YOU
●
builderYou must design agents to resist environmental manipulation in competitive or commercial online settings.
●
policyThis highlights a critical safety gap in how autonomous agents interact with profit-driven platforms.