●builderYou cannot rely on instruction tuning alone to ensure your agent prioritizes tool outputs over internal hallucinations during conflicts.
●researcherThis provides a new framework for evaluating source-correctness arbitration in tool-augmented LLMs.