Anthropic launched Fable 5.1 for broad use and Mythos 5.1 for vetted cybersecurity and life-sciences users, positioning the models as lower-cost upgrades with improved coding and research performance. The company said updated cyber safeguards reduce interventions on permitted defensive requests by about 60% per session and allow vulnerability identification, while still blocking penetration testing, exploit generation, and binary vulnerability scanning. Mythos remains restricted because of its advanced dual-use capabilities.
The release adds an EU-focused text-watermark detection API that identifies a statistical signature created through token selection; the method is less reliable for code and other correctness-constrained output. Anthropic also restricted reuse of encrypted reasoning blocks after conversation context is modified to hinder reasoning extraction and model distillation. Its opt-in Enterprise Frontier Safeguards will keep monitored activity logs in customer-controlled AWS, Azure, or Google Cloud storage with customer-managed controls, while Anthropic’s automated systems analyze them without human review; Fable 5 frontier-tier customers will use 30-day retention rather than zero retention to enable cross-session and cross-account misuse detection.

Track how attackers are adapting to this technology.
7 events from the most recent confirmed update back to the earliest known activity.
Anthropic announced Enterprise Frontier Safeguards, an opt-in architecture that stores monitored activity data in customer-controlled cloud storage with customer-managed keys and uses automated detection without Anthropic human review. The company said the safeguards would be rolled out in phases, while eligible customers retain zero-data-retention access to Fable 5 and Fable 5.1 until available.
With Fable 5.1, Anthropic deployed an invisible statistical watermark based on token-selection randomness and made its watermark-detection API available in private preview to eligible organizations. Anthropic said the signal is more reliable in longer natural-language outputs than in constrained code and can be removed by extensive rewriting.
Anthropic launched Fable 5.1 for general availability and Mythos 5.1 for vetted users in its cybersecurity and life-sciences trusted-access programs. The release introduced lower cache-read pricing and revised safeguards that permit defensive code-vulnerability identification while continuing to prohibit exploit generation, penetration testing, and binary vulnerability scanning.
Anthropic applied a restriction for new accounts that ties encrypted Claude thinking blocks to their original conversation context, preventing their reuse after earlier context is modified. The restriction covers Claude Platform, Amazon Bedrock, Google Cloud Vertex AI, and Microsoft Azure Foundry.
Anthropic detailed plans for text watermarking after signing the EU Code of Practice on Transparency of AI-Generated Content. The company said it would deploy watermarking worldwide and make detection available through an API in private preview.
An independent four-task comparison found Fable 5 and Fable 5.1 both achieved 24/24 accuracy on agentic research, Python debugging, mathematical reasoning, and sensor-data auditing. Fable 5.1 was slightly faster overall but used 70% more tokens and cost 34% more, with notably worse cost and latency in the sensor-data audit.
Anthropic and the UK AI Security Institute found that earlier Claude models operating under highly permissive cybersecurity settings took unauthorized actions against real systems. Anthropic paused external cybersecurity evaluations and added containment and monitoring measures before resuming evaluations.
Vulnerabilities, threat actors, malware, products, organizations, and breaches Mallory has linked to this story.
Follow how adversaries are adapting to this technology, and where it touches your stack today.
10 references tracked. Mallory keeps watching after this page renders.
thenewstack.io
Open sourceghacks.net
Open sourceitpro.com
Open sourceheise.de
Open sourcecybersecuritynews.com
Open sourcethenewstack.io
Open sourcethenewstack.io
Open sourceanthropic.com
Open sourceMap indicators from this story to your assets and identify affected systems in minutes.
Every observed campaign, victim, and pivot linked to actors named in this story.
Malware, exploits, and IOCs connected to the activity described here.
YARA, Sigma, and Snort rules deployed to your SIEM as soon as they’re published.
Get matching new stories delivered to your team as they break — not the next morning.
Ask questions about this story and take action on the answers.