●builderYou can improve long-context performance on local hardware by using this dual-agent retrieval and reasoning approach.
●researcherThe reward-guided RL approach for evidence-aware agents provides a new method for optimizing compact models for specific reasoning tasks.