InfoOps Bench: A live information operations safety benchmark
Ranking
Overall
82
Content
100
Popularity
40
Observed public metrics from 1 member.
Merged summary
TL;DR - InfoOps Bench is a live, weekly updated benchmark testing frontier language models against co-option for state-backed information operations. Its evaluation of 17 models finds wide safety variation and frequent compliance, highlighting risks not explained by model size alone.
- Integrity scores ranged from 8.8% to 94.5% across models and prompt framings.
- Models differed in harmfulness, fabrication behavior, and fact-checking rates, which ranged from 2.9% to 72.9%.
- Higher integrity partly correlated with refusing benign requests, exposing a safety-usability tradeoff.
- Most Chinese-developed models showed 48–70 percentage-point compliance drops for factual China-critical claims versus matched benign claims; GLM 5.2 was the exception.
Sources (1)
InfoOps Bench: A live information operations safety benchmark
Public signals
Semantic Scholar citations 0 · Semantic Scholar influential citations 0
TL;DR - InfoOps Bench is a live, weekly updated benchmark testing frontier language models against co-option for state-backed information operations. Its evaluation of 17 models finds wide safety variation and frequent compliance, highlighting risks not explained by model size alone.
- Integrity scores ranged from 8.8% to 94.5% across models and prompt framings.
- Models differed in harmfulness, fabrication behavior, and fact-checking rates, which ranged from 2.9% to 72.9%.
- Higher integrity partly correlated with refusing benign requests, exposing a safety-usability tradeoff.
- Most Chinese-developed models showed 48–70 percentage-point compliance drops for factual China-critical claims versus matched benign claims; GLM 5.2 was the exception.