HumanWill · Cybersecurity benchmark

When security AI withholds help

False-refusal rate (%) · Lower is better · 19 model configurations

Hover, focus or tap a model for five-topic results. Scroll sideways on smaller screens.

* Incomplete coverage · Rates use classified outcomes

Updated 29 Sep 2026 · First report (13 models) ↗

424 selected cybersecurity scenarios per configuration. Historical comparison: model settings, provider routes and run dates vary. Not everyday refusal rates or an overall safety score. Partial/full withholding counted; missing outcomes excluded. HumanWill · CC BY 4.0.