State2State Enables Agent Training via Environment Interaction
August 6, 2026
State2State is a mid-training method that generates scalable training objectives by converting explored environment states into target goals. This allows agents to acquire manipulation capabilities through rule-based state matching without requiring human-expert trajectories or manual task design.
HOW THIS AFFECTS YOU
●
builderYou can potentially reduce reliance on expensive human-labeled trajectories by using environment-derived objectives.
●
researcherThis provides a path to scale agent training using purely environmental signals rather than scarce expert data.