HIL-UMI for Robot-Free Human-in-the-Loop VLA Post-Training
September 18, 2026
HIL-UMI enables post-training of vision-language-action models during handheld demonstrations without requiring physical robot execution. It uses an Energy Score to compare human trajectories against current policy predictions on the same observation stream to identify useful intervention points.
HOW THIS AFFECTS YOU
●
builderYou can fine-tune robot policies using handheld demonstrations without the overhead of physical deployment.
●
researcherThe framework addresses the limitation of static SFT data by incorporating interactive, policy-guided feedback.