InfoOps Bench: Evaluating LLM Susceptibility to State-Backed Information Operations
July 31, 2026
A live benchmark tracks 2,100 information operations from Russian, Chinese, and Iranian assets to measure frontier model integrity. Testing across 17 models shows integrity scores—defined by refusal rates—varying widely from 8.8% to 94.5% regardless of model size.
HOW THIS AFFECTS YOU
●
researcherYou can use this dynamic benchmark to evaluate how easily your models can be co-opted for misinformation.
●
policyThis provides a quantitative metric for assessing the geopolitical safety and integrity of deployed LLMs.