Efficient Training Recipes for Looped Language Models
October 2, 2026
A new training pipeline reduces the compute budget for looped language models from 7.7T to 310B tokens while maintaining reasoning performance. The 1.4B LoopLM outperformed parameter-matched dense models across 12 benchmarks using high-quality mid-training and exit-gate regularization.
HOW THIS AFFECTS YOU
●
builderYou can deploy more efficient recurrent models that offer higher reasoning density per parameter.
●
researcherYou can leverage these recipes to study recurrent architectures without massive multi-stage budgets.