AnySearch Framework Enables Budget-Aware LLM Tool Search via Reinforcement Learning
September 2, 2026
AnySearch uses curriculum reinforcement learning and budget state injection to train a single policy that adapts to varying inference constraints. The framework optimizes a composite reward to balance answer accuracy with computational efficiency during autonomous search.
HOW THIS AFFECTS YOU
●
builderYou can deploy agents that dynamically adjust their search depth based on real-time compute or cost constraints.
●
researcherThe two-phase training scaffold offers a new way to internalize resource constraints into model policies.