[arXiv]score: 0.14
Self-Referential Induction Increases Response Instability Relative to Unresolvable and Verifiable Questions in Large Language Models
August 14, 2026
Self-referential prompts induce higher response instability in LLMs compared to unresolvable or verifiable questions. Testing via Gemini API at 0.7 temperature shows self-referential claims reach a mean pairwise cosine similarity instability of 0.343 +/- 0.047, measured by sentence embeddings of extracted core claims across 30 independent trials.
DAILY DIGEST
you don't check 9 sources — we do. one email every morning, read in 2 min. free. unsubscribe anytime. privacy