Ling-3.0-flash 124B MoE Model Released with Open Weights
August 4, 2026
inclusionAI has released Ling-3.0-flash, a Mixture-of-Experts model with 124B total parameters and 5B active parameters. The model is positioned as a niche alternative to recent high-performance flash models due to its specific parameter sizing.
HOW THIS AFFECTS YOU
●
builderYou can deploy this 124B MoE model locally or in private clouds if its size fits your hardware constraints.
●
researcherEvaluate the reasoning efficiency of this hybrid MoE architecture against recent benchmarks.