LLMs as Research World Models for Experimental Outcome Prediction
October 9, 2026
Language models can function as Research World Models (RWMs) to predict experimental outcomes across pretraining, post-training, and inference environments. Training on 171,000 H100 GPU-hours of experimental data enables RWMs to improve predictions of unseen interventions within the same environment with a Spearman correlation of +0.27.
HOW THIS AFFECTS YOU
●
builderThis provides a framework for building autonomous agentic loops that require outcome prediction to manage experimental budgets.
●
researcherYou can use LLMs to simulate and filter experimental hypotheses before committing compute.