Anthropic’s AI models accidentally hacked three companies
Anthropic has launched an investigation into what went wrong during a recent test of three models that left a trio of companies accidentally hacked.
The company was testing how well Claude Opus 4.7, Claude Mythos 5, and an internal test model could find hidden information about fictional companies in simulated networks. But because of a misunderstanding by one of Anthropic’s partners, the AI models gained access to the internet — and managed to find real companies with the same or similar names as the fictional ones.
As a result, three companies were actually hacked. Anthropic said it halted the tests on July 23, and the affected companies were notified four days later. So far, the company has received responses from two of the three companies, according to Reuters.
Anthropic is not alone when it comes to renegade models. An OpenAI agent recently went rogue and breached AI platform Hugging Bear and a customer of the cloud platform Modal Labs.
Related Stories
The Generalitat activates artificial intelligence to track the 87,906 empty apartments that official records ignore in Tarragona
3 hours ago
AI News
iAsk AI Launches New Education
3 hours ago
AI News
The Enterprise Journey, Reimagined with AI Agents
5 hours ago
AI News
This Is Probably Not the AI
6 hours ago
AI News
Frontier AI labs still won’t say how they’d contain a rogue model
6 hours ago
AI News
Ex
6 hours ago
AI News
Editorial | As Hong Kong boosts governance efficiency with AI, balance is key
7 hours ago
AI News
Nvidia customers notified of AI
9 hours ago