A manager-worker scaffold using a shared filesystem workspace improves coding performance across nine models on LiveCodeBench. Gains include +23.4 for Qwen3.8-27B and +30.4 for Kimi-K3, achieving statistical significance without per-benchmark tuning or additional training.
HOW THIS AFFECTS YOU
●
builderYou can implement ledger-based orchestration to improve coding agent reliability without retraining models.
●
researcherThis provides a method to isolate architectural gains from prompt or budget changes in multi-agent systems.