[arXiv]score: 0.19
Clinician use of language models diverges from how the models are evaluated
October 9, 2026
Analysis of 127,833 clinical queries reveals that real-world utility diverges from standard examination-based benchmarks. While benchmarks focus on medical knowledge, 65% of actual clinician use involves documentation, administration, and knowledge retrieval, with only 3.7% of queries directed toward diagnostic tasks.
DAILY DIGEST
you don't check 9 sources — we do. one email every morning, read in 2 min. free. unsubscribe anytime. privacy