Anthropic Expands Access to Latest AI Models for Cyber Firms
Anthropic PBC is expanding access to its most advanced AI models, allowing a select group of organizations to test the startup’s cutting-edge cyber capabilities in collaboration with the US government.
The San Francisco-based company will grant verified organizations access to its most capable models, including Claude Opus 5.5, Claude Sonnet 5.5, Claude Mythos 5.1, and new models moving forward, according to a statement Tuesday. Those with access can carry out “high risk offensive testing” of the safety systems designed to protect critical infrastructure like power grids, banks and flight operating systems, it said.
Anthropic granted US government agencies, financial institutions, and major software providers limited access to its Mythos model in April under Project Glasswing. The Mythos release was a watershed moment for cybersecurity because of the model’s ability to partially automate the work of sophisticated hackers. Since then, a spate of incidents in which Anthropic and OpenAI’s models inadvertently hacked governments, companies and non-profit organizations has raised concerns about the safety of the technology.
Anthropic said Tuesday the latest limited release will include every member of Project Glasswing and require a review of all new organizations in partnership with the US government. In August, Washington said it would partner with private-sector firms to launch cyber attacks abroad, expanding the scope of national security operations that until now have been largely conducted by government agencies.
On Tuesday, JPMorgan Chase & Co. Chief Executive Officer Jamie Dimon told Bloomberg Television that Mythos had boosted global cybersecurity risks “10-fold.”
Read More: Dimon Says Anthropic’s Mythos Pushed Cyber Risk Up 10-Fold
Different types of cybersecurity teams will be allowed to apply for lower levels of access. Red teams, which conduct authorized hacking of targets in order to find and fix weaknesses, will have more permissions to security test AI models, including for offensive testing, but Anthropic said it will still block behavior that can cause physical harm and mass disruption. Verified defense cybersecurity teams will have a slightly lower level of access, but it will allow tasks such as reverse engineering malware and incident response.