Anthropic restored global access to its Fable 5 model after temporarily restricting it in response to US government concerns that a jailbreak could enable misuse of the model’s cybersecurity capabilities. The company said Amazon researchers identified a safeguard bypass, prompting Anthropic to redeploy Fable 5 with new classifiers designed to block more cyber-related tasks while reducing false positives. Anthropic also resumed access to Mythos 5 for a limited set of US organizations and said testing found other, less capable models could identify the same vulnerabilities, suggesting the bypass did not unlock uniquely advanced offensive capability.
The episode has intensified scrutiny of how governments and AI vendors assess cyber risk from frontier models. Orion Policy Institute argued that Washington’s temporary ban and reversal exposed a broader weakness in real-world cyber-risk measurement, especially when policymakers must judge whether a model meaningfully increases offensive capability or merely replicates what existing tools can already do. Anthropic said it is now working with the US government and partners including Amazon, Microsoft, Google, and Glasswing participants to build a framework for evaluating AI jailbreak severity and determining appropriate response measures.

Mallory correlates global threat intelligence with your attack surface — know if you’re exposed before adversaries strike.
3 events from the most recent confirmed update back to the earliest known activity.
Anthropic said access to Mythos 5 had also been restored for a set of US organizations. The company additionally said it was working with the US government and partners including Amazon, Microsoft, Google, and Glasswing participants on a framework for assessing AI jailbreak severity and response.
Anthropic temporarily restricted access to its Fable 5 model after US government concerns about a jailbreak that could enable misuse of the model's cybersecurity capabilities. The company said the issue involved a safeguard bypass method identified by Amazon researchers.
On 2026-07-01, Anthropic restored global availability of Fable 5 after updating the model with new classifiers intended to block more cybersecurity tasks while reducing false positives. Anthropic also said testing showed less capable models could identify the same vulnerabilities and that the technique did not expose unique "Mythos-level" cyber capabilities.
Vulnerabilities, threat actors, malware, products, organizations, and breaches Mallory has linked to this story.
2 references tracked. Mallory keeps watching after this page renders.
orionpolicy.org
Open sourcepcmag.com
Open sourceMap indicators from this story to your assets and identify affected systems in minutes.
Every observed campaign, victim, and pivot linked to actors named in this story.
Malware, exploits, and IOCs connected to the activity described here.
YARA, Sigma, and Snort rules deployed to your SIEM as soon as they’re published.
Get matching new stories delivered to your team as they break — not the next morning.
Ask questions about this story and take action on the answers.