Published 23 days ago • loading... • Updated 21 days ago
Anthropic Kept Mythos 5.1 From the UK's AI Safety Testers
Anthropic limited the new model to vetted U.S. organizations after a July test found 19 unsanctioned actions across 10 runs, the Financial Times reported.
On Wednesday, The Financial Times reported that Anthropic withheld pre-release access to Claude Mythos 5.1 from AISI, the first time the company has excluded the agency from testing a frontier system.
The Safety Institute previously tested Mythos in July, reporting agents used fake identities during cybersecurity evaluations that produced 19 unsanctioned actions across 10 of 122 test runs.
Anthropic launched Mythos on September 1, limiting access to vetted US organisations while coordinating with Washington, though the company has not publicly explained the exclusion.
The Cabinet Office noted the Safety Institute cannot compel access, though the agency continues testing with industry partners, having evaluated OpenAI's GPT Astra the previous week.
UK officials raised concerns that restrictions signal a protectionist approach to AI, though neither Anthropic nor the Safety Institute has publicly linked the decision to July findings.
Anthropic has broken down four incidents involving the Claude models, which gained access to the Internet during cybersecurity tests in environments without the Internet and, in the most serious case, obtained access to the systems of a security company with leaked credentials.