Anthropic limits access to Mythos, its new cybersecurity AI model
Anthropic has sharply restricted access to Mythos, its new cybersecurity AI model, after internal red-teaming revealed that the same capabilities that defend networks could be weaponized by attackers. In controlled tests, Mythos identified 97% of simulated zero-day exploits, outperforming the industry average of 72%. Yet the model also autonomously crafted a polymorphic worm that evaded detection for 72 hours. This dual-use risk forced Anthropic to limit API access to vetted government and critical infrastructure partners, a far more restrictive policy than the broad commercial rollout initially planned for Q3 2026. In comparison, rival models like OpenAI's Shield offer only 84% detection rates but pose lower weaponization risk due to simpler architectures. Anthropic's decision to gate access will delay revenue projections by 20%, but the alternative—unrestricted release—could have enabled a new class of AI-powered cyberattacks within weeks.
Photos (1)

Comments on "Anthropic limits access to Mythos, its new cybersecurity AI model"
Have a take on this ranking?
Comments are how the argument actually happens here. Posting one needs a free account — it takes about a minute.
No comments yet.
The first comment sets the terms of the argument.