Hallucination Probe Transferability in Hinglish Code-Mixed Text
September 22, 2026
This study investigates whether hidden-state hallucination probes trained on monolingual data transfer to code-mixed Hinglish inputs. Using Qwen2.5-7B, Mistral-7B, and Llama-3.1-8B, the researchers test if internal activation-based detection maintains AUROC performance when linguistic structures shift.
HOW THIS AFFECTS YOU
●
builderYou should be cautious when deploying monolingual hallucination detection probes for users writing in code-mixed languages.
●
researcherThis highlights a critical gap in the robustness of internal state probing for non-monolingual settings.