PhysBrain 1.5: Unified Physical Foundation Model for Robotics
September 13, 2026
PhysBrain 1.5 integrates vision-language understanding with action generation and future state prediction. It uses a common learning framework to jointly optimize language, end-effector motion, and visual targets via autoregressive next-token prediction trained on human interaction videos.
HOW THIS AFFECTS YOU
●
builderYou can use a single model for both environmental understanding and robot motion control.
●
researcherThis demonstrates the effectiveness of treating physical interaction as a unified sequence prediction task.