GuardianBench Evaluates Latent Contextual Risk in Embodied AI Models
August 25, 2026
GuardianBench introduces 3,024 instruction-scene contrastive pairs to isolate safety risks in embodied AI where benign instructions become hazardous in specific visual contexts. Current vision-language models show poor performance, with an average pair accuracy of only 24.1% in identifying these latent risks.
HOW THIS AFFECTS YOU
●
builderYou must implement stricter instruction-scene binding to ensure safety in embodied robotics.
●
researcherThis benchmark provides a new axis for evaluating safety beyond visual context alone.