JetBrains analysis finds CaveMan and RTK token savers fail to meet claims
July 22, 2026
JetBrains testing of Claude 3.5 Sonnet shows CaveMan achieves only 8.5% token savings against an advertised 65%. RTK was measured to be 7.6% more expensive than baseline at low reasoning efforts, despite claims of 60–90% savings.
HOW THIS AFFECTS YOU
●
builderDo not rely on advertised token-saving prompts for cost modeling in production agent workflows.
●
founderBe wary of integrating third-party efficiency tools that lack verified performance metrics on high-reasoning models.