ContextProgress-Bench Evaluates Progress Reward Models in Long-Horizon Tasks
September 30, 2026
Researchers introduce ContextProgress-Bench, a benchmark featuring 24 manipulation tasks designed to test Progress Reward Models (PRMs) on context-dependent progress estimation. The study highlights how current PRMs fail when progress requires information from previous frames rather than just the current observation.
HOW THIS AFFECTS YOU
●
researcherYou should account for temporal context and state recall when training reward models for long-horizon embodied agents.