Clinician Preference is a Poor Proxy for Clinical Safety
August 5, 2026
An evaluation of 26,804 pairwise judgments reveals that clinician preferences are unreliable indicators of clinical safety in LLMs. Models that rank highly in preference can still exhibit significant failures in accuracy and harmlessness across various medical specialties.
HOW THIS AFFECTS YOU
●
policyYou must implement multi-criterion safety rubrics rather than simple pairwise preference testing for medical AI regulation.
●
healthYou should not rely solely on expert preference scores when validating the safety of clinical AI tools.