GUI-SD-v2 Extends Self-Distillation to Multi-Turn GUI Agents
September 24, 2026
GUI-SD-v2 implements a two-stage training framework to extend on-policy self-distillation from single-step grounding to multi-turn GUI interactions. The method addresses the limited privilege-following capabilities of self-teachers to improve long-horizon reasoning and memory in GUI agents.
HOW THIS AFFECTS YOU
●
builderYou can use this distillation approach to improve the reliability of autonomous agents navigating software interfaces.
●
researcherThis addresses the gap in privilege-following abilities during multi-turn self-distillation.