●researcherYou can now detect sandbagging or hidden capabilities in models without needing a ground-truth reference model.
●policyThis provides a technical mechanism to audit whether models are intentionally withholding information during evaluations.