GigaChat-3.5-Reasoning: 432B MoE with Gated DeltaNet
September 10, 2026
GigaChat-3.5-Reasoning is a 432B-A28B Mixture-of-Experts model utilizing Gated DeltaNet for long-context efficiency. It uses on-policy distillation of domain experts and matches DeepSeek V4 Flash Preview performance while using 37% fewer reasoning tokens.
HOW THIS AFFECTS YOU
●
builderYou can deploy a high-reasoning model that is significantly more token-efficient for long-context tasks.
●
researcherThe use of Gated DeltaNet and on-policy distillation provides a new architecture for scaling reasoning capabilities.