●builderYou can deploy this model to significantly reduce inference latency and cost for reasoning-heavy tasks.
●researcherThis demonstrates the effectiveness of on-policy distillation for optimizing reasoning-length efficiency without sacrificing accuracy.