Unauthorized Access to a "Dangerous" AI
In a significant security lapse, a small group of users operating within a private Discord channel reportedly gained unauthorized access to Anthropic's Mythos AI model, a technology the company itself described as capable of enabling "dangerous cyberattacks." This breach reportedly occurred on the very day Anthropic publicly announced its limited release of Mythos as part of Project Glasswing, an initiative designed to provide select companies and government users with access to the model for defensive cybersecurity purposes. Anthropic has confirmed it is investigating the report, stating that the unauthorized access appears to have been facilitated "through one of our third-party vendor environments."
The Mythos AI model is touted for its ability to identify and exploit vulnerabilities across major operating systems and web browsers, with Anthropic claiming it can surpass most skilled humans in finding and exploiting software flaws. Concerns about the model's potential for misuse led Anthropic to restrict its release, sharing it only with a limited group of major companies including Amazon, Apple, Cisco, JPMorgan Chase, and Nvidia. The incident has ignited fresh concerns regarding the control and security of high-end cybersecurity tools, especially those with dual-use risks.