H3-World Enables Language-Native Control for Video and Gameplay
September 2, 2026
H3-World maps textual instructions to character and camera actions by injecting prompts into a pretrained text pathway. The method achieves controllable motion using only 8,000 gameplay samples and 0.199% trainable parameters via LoRA.
HOW THIS AFFECTS YOU
●
researcherYou can utilize this highly parameter-efficient fine-tuning approach for temporal grounding in video models.
●
designerYou can use text-based instructions to control complex camera and character movements in generative environments.