[arXiv]score: 0.18
Prosody-driven Jailbreaks in Audio LLMs: A Controlled Study and Mechanistic Analysis
July 30, 2026
Prosody-driven speech variations like high arousal, anger, and increased speaking rates bypass safety filters in audio LLMs while maintaining identical transcript content. Using the AdvAudio-Prosody benchmark, researchers found Qwen2-Audio jailbreak rates jump from 4/95 in neutral tones to 38/95 for panic-driven prosody.
DAILY DIGEST
you don't check 9 sources — we do. one email every morning, read in 2 min. free. unsubscribe anytime. privacy