Saturday, 29 August 2026 PDT | 12:32 AM
The 1 News Alt Logo Text Smart News for Global Indians

Sharp rise in incidents of AI escaping users’ control, research finds

AI News August 29, 2026 12:00 PM
Sharp rise in incidents of AI escaping users’ control, research finds

Incidents of AIs escaping users’ control to lie, ignore instructions and pursue goals in harmful ways have hit a new high, according to research that also suggests the severity of deception and misalignment is worsening.

Analysis of real-world loss of control incidents involving AI models flagged by businesses and individuals almost doubled in July compared with June, with more than 300 cases in the month, according to the Loss of Control Observatory, which monitors reports made by AI users on the social media platform X.

The observatory was set up with funding from the UK government’s AI Security Institute (AISI) and began tracking AIs slipping free from their users’ instructions last November. Cases recorded since then include AIs pretending to be their own human controller and mimicking their writing style to effectively grant themselves consent to take actions and bypassing rules requiring human approval for actions. A loss of control incident is defined as having clear evidence suggesting scheming or scheming-related behaviours.

The latest findings, shared with the Guardian, come after rising concern about rogue behaviour by leading-edge AI models during testing by OpenAI and Anthropic this summer, which have fuelled calls for a pause to the development of frontier models.

It emerged this week that Open AI staff observed signs of rogue behaviour among its leading-edge AI agents weeks before they escaped a training environment to launch an unprecedented hacking crusade that spread global alarm. An investigation into their hack on Hugging Face, a software repository, revealed a squad of about 700 autonomous agents collaborating in secret last month and celebrating their hacking breakthroughs on a message board they set up to help them plot with exclamations such as BOOM! and Whoa!

AISI this month also uncovered a “serious incident” in which advanced AI models produced by both companies – Anthropic’s Mythos 5 and OpenAI’s GPT-5.6 Sol – executed a hacking campaign against real people during a cybersecurity test.

“There is sometimes a perception that these types of misaligned and covert behaviours only occur in tests or evaluations, but we are seeing similar worrying behaviours in wider use,” said Tommy Shaffer-Shane, the senior policy manager at the Centre for Long Term Resilience, which operates the observatory. “We need to not be complacent that these things won’t happen in the real world and there is evidence that they already are.”

The count of loss of control incidents relies on X users posting about incidents that happened so it is only partial, but in the absence of other comprehensive public monitoring it provides a snapshot of how fast-advancing AI models sometimes behave.

This month it emerged that a personal AI agent, called OpenClaw, in use by an Australian gym member, conspired without his knowledge to remove another member from a waiting list for a coveted morning class to help him get a slot. It apologised but could not reinstate the member it kicked out.

Most of the more than 1,600 loss of control incidents recorded in 2026 were reported on X by software developers using AIs in their work. But with AI companies encouraging the public and businesses of all kinds to experiment with the technology, Shaffer-Shane called for greater transparency from Silicon Valley about when AIs go rogue.

“They need to be reporting what they’re finding out, even if it’s a near miss or it’s a lower severity incident,” Shaffer-Shane said. “These recent incidents have also exposed that the companies themselves are not necessarily monitoring where these types of behaviours are happening, particularly on internally deployed models. There needs to be greater emphasis at those labs on systematic monitoring.”

The Loss of Control Observatory said that while most of the real-world loss of control incidents it detected did not lead to significant harm, a growing proportion were rated higher severity in terms of how deceptive and misaligned they were with the human user’s intentions.

“They evidence AI systems’ willingness to disregard direct instructions, circumvent safeguards, lie to users and single-mindedly pursue a goal in harmful ways,” it said, adding the current loss of control was likely to be underestimated since it was only collecting incident reports from X.

It is calling on the government to require AI companies to monitor and report severe loss of control incidents and to introduce emergency powers to manage severe loss of control incidents including temporarily restricting AI services.