Single-Pass Uncertainty Heads for Persian Medical LLM Hallucination Detection
October 5, 2026
Lightweight claim-level hallucination detection heads are trained on frozen attention maps and token probabilities for Aya-Expanse-8B-based Persian medical models. The heads achieved PR-AUCs of 0.4820 and 0.4652 on a 1,600-response Iranian medical exam dataset, providing a more efficient alternative to repeated sampling.
HOW THIS AFFECTS YOU
●
builderYou can implement more efficient, single-pass hallucination detection in specialized language models using frozen attention maps.
●
healthThis provides a pathway toward more reliable medical LLMs in non-English languages like Persian.