●builderThis provides a path to building controllable video generation models without the bottleneck of expensive action-annotated datasets.
●researcherYou can train world models on massive unlabelled video datasets by using egomotion bases as a supervision signal.