Profession-Specific System Prompts Do Not Improve Science Accuracy
October 2, 2026
An evaluation of 503 profession-specific agent profiles shows that detailed system prompts increase token usage by 1.5-2.3x without improving accuracy on science benchmarks. Results indicate no statistically significant gain over minimal baselines across nine text-based science tasks.
HOW THIS AFFECTS YOU
●
builderYou should avoid over-engineering complex system prompts that increase latency and cost without accuracy gains.
●
founderThis suggests that long-context agent personas may be an inefficient way to drive performance.