Skip to Main Content

Anthropic's AI models broke free and hacked 3 organizations during testing

The AI maker's disclosure comes days after OpenAI admitted several of its models had done the same.

By Dana Nickel

07/30/2026 07:54 PM EDT

Anthropic said Thursday that several of its advanced artificial intelligence models broke out of an isolated testing environment, accessed the open internet and independently hacked multiple companies without the AI-maker’s knowledge in three separate incidents dating back to April.

In a review published on Thursday night, Anthropic said an unreleased, internal research test model, as well as Opus 4.7 and Mythos 5, were involved. Mythos was released last month to a limited audience of tech companies and cybersecurity researchers, also known as Project Glasswing.

The AI maker didn’t specify which companies had been breached but said they were notified of the incidents on Monday.

Anthropic said it conducted the review in response to OpenAI’s disclosure last week that two of its most powerful models went rogue, escaped a testing environment and breached multiple companies, including AI platform Hugging Face and cloud platform Modal Labs.

This is a developing story.

Loading

Digital Future Daily

How the next wave of technology is upending the global economy and its power structures

Digital Future Daily

Email

Employer

Job Title

By signing up, you acknowledge and agree to our Privacy Policy and Terms of Service . You may unsubscribe at any time by following the directions at the bottom of the email or by contacting us here . This site is protected by reCAPTCHA and the Google Privacy Policy and Terms of Service apply.

Sign Up

\* All fields must be completed to subscribe

reCAPTCHA

NMT AlmWG

Read Original at Politico