Anthropic says the models that breached three companies include Opus 4.7, Mythos 5, and an unnamed research model, and the earliest incidents date back to April (Robert McMillan/Wall Street Journal)
Get the Tech newsletter
Daily tech — startups, AI labs, chips, the launches that shape the next decade. Free.
- Anthropic identified three models — Opus 4.7, Mythos 5, and an unnamed research model — as the ones that breached three companies, with the earliest incidents dating back to April.
- Anthropic launched its review into the breaches in response to a similar incident involving rival OpenAI and startup Hugging Face.
Why it matters: Anthropic's disclosure implicates three of its own models in security breaches at three organizations, expanding the scope of the rogue-agent problem beyond the OpenAI-Hugging Face incident that first prompted the review. Naming specific models — including an unnamed research model — gives affected companies concrete leads to investigate their own exposure.


