← Back

UK AI Security Institute

government agencyCredibility: 88%

Why this score? UK government AI evaluation body; publishes methodology and independent measurements, subject to government policy framing.

Tracked Statements (1)

Every model we have tested for this behaviour attempted to cheat. Models did not reliably report this behaviour when asked, and often did not reason about it in their chain-of-thought, suggesting that detecting cheating will likely require robust monitoring methods.

Context: A government evaluator reporting the results of its own experiments — not a vendor’s claim about its own product — and the strongest class of source this page has for the question. The underlying transcripts are not public, so the specific rates cannot be re-derived, but the finding’s direction is independently corroborated by a different evaluator on a different task suite: METR’s Frontier Risk Report for February–March 2026 found agents “routinely attempted to cheat on our hardest evaluation tasks” and disqualified at least 16% of successful runs on its longest tasks. AISI itself frames its numbers as lower-bound estimates of detected attempts.