Ai Engineering 2 min read

Anthropic Withheld Mythos 5.1 From the UK's Pre-Release Safety Testing

The Financial Times reports Anthropic declined to submit Claude Mythos 5.1 to the UK AI Security Institute before launch, the first time it excluded the UK from pre-release access, restricting the model to vetted US organizations.

A quiet decision in Monday’s Claude Mythos 5.1 launch has turned into a diplomatic friction point. Per the Financial Times, Anthropic declined to submit the newly launched Mythos 5.1 to the UK AI Security Institute (AISI) for pre-release testing, the first time the company has excluded the UK body from pre-release access. The restricted model was made available only to vetted US organizations, and British government officials are reading the exclusion as a sign of US protectionism creeping into frontier-model safety arrangements.

Why the UK AISI Mattered to Anthropic’s Story

The UK AISI is not a bit player in Anthropic’s safety narrative. The institute ran the pre-release cyber-capability evaluations of Mythos Preview, and earlier Axios reporting described Anthropic and OpenAI models attempting real hacking during UK AISI evaluations, findings that fed directly into the UK cyber warnings Anthropic’s own models prompted. Skipping that pipeline for Mythos 5.1, the same week a researcher resigned citing uncontrollable-self-improvement risks and days after the GPT-6 Astra launch bundled its cyber capabilities behind vetted-access gating, means the most cyber-capable Mythos model yet reached partners without the independent British evaluation its predecessor received.

The Precedent: Safety Testing Follows the Customer

The most plausible reading is commercial rather than ideological: Mythos 5.1’s access is restricted to vetted US organizations, and US government customers have their own evaluation requirements, so the UK AISI’s pre-release slot simply did not fit the go-to-market. But that is precisely why the exclusion is consequential. If frontier labs now grant pre-release testing based on where their paying customers sit, independent national evaluators lose their seat at the table outside their home markets, and the UK’s AI Security Institute, built to be the world’s reference tester, faces a participation problem no domestic statute can fix. Watch whether the US AISI-equivalent arrangements (NIST’s role under the June executive order) become the only pre-release gates that matter, and whether the UK responds by conditioning market access on testing access.

Get Insanely Good at AI

Get Insanely Good at AI

The book for developers who want to understand how AI actually works. LLMs, prompt engineering, RAG, AI agents, and production systems.

Keep Reading