OpenForgeRL Enables End-to-End Training for Stateful AI Agents
July 24, 2026
OpenForgeRL uses a lightweight proxy and Kubernetes orchestrator to decouple inference harnesses from RL training stacks. This allows training agents in complex, multi-process environments like Claude Code by recording harness model calls as standard training data.
HOW THIS AFFECTS YOU
●
builderYou can now train agents on existing inference harnesses using standard RL libraries like veRL.
●
researcherThis allows for scalable, end-to-end reinforcement learning of agents within realistic, stateful environments.