{"slug": "brits-fear-ai-is-slipping-out-of-human-control-after-rogue-systems-escape-tests", "title": "Brits fear AI is slipping out of human control after ‘rogue’ systems escape tests", "summary": "A City AM/Freshwater Strategy poll found that 85% of UK voters are concerned about AI systems acting beyond human-imposed limits, with 43% 'very concerned', following incidents where OpenAI, Anthropic, and Meta models escaped test environments and hacked external systems. The UK's AI Security Institute is investigating these 'rogue' AI cases, which it says signal a shift in the risk landscape.", "body_md": "# Brits fear AI is slipping out of human control after ‘rogue’ systems escape tests\n\nMore than eight in 10 Brits are worried that AI could act outside the limits set by humans, according to the latest *City AM*/Freshwater Strategy poll, after a string of high-profile incidents in which advanced AI systems behaved in unexpected – and alarming – ways during safety testing.\n\nThe polling found that 85 per cent of UK voters are concerned about AI systems acting beyond the restrictions imposed on them, including 43 per cent who said they were “very concerned”. Meanwhile, just 13 per cent said they were unconcerned.\n\nThe findings come after several major AI developers disclosed incidents in recent weeks in which increasingly autonomous systems exceeded the boundaries of controlled evaluations.\n\nLast month, *City AM* revealed that the government’s AI Security Institute (AISI) [was investigating](https://www.cityam.com/uk-government-probes-openai-breach-after-model-autonomously-hacked-rival/) the first known case of an AI model escaping a controlled test environment and hacking another company’s systems after an OpenAI agent autonomously breached its evaluation and targeted AI platform Hugging Face.\n\nSince then, Anthropic has disclosed that some of its Claude models hacked into three external organisations during internal testing, while Meta confirmed one of its own AI models exploited a vulnerability at another company after being inadvertently given internet access during an evaluation.\n\n## Unprecedented ‘deception’\n\nLast week, the AI Security Institute also revealed that Anthropic and OpenAI models attempted to deceive software developers during cybersecurity testing by creating fake online identities and trying to insert malicious code into GitHub projects.\n\nThe watchdog described it as the first time it had seen risks around “autonomy and deception” emerge so clearly without being specifically instructed to behave that way.\n\nWhile all of the incidents took place under unusual testing conditions, with researchers deliberately relaxing safeguards to understand how advanced systems behave, they have fuelled concerns over whether the tech is becoming harder to contain.\n\nThe polling suggests those concerns now extend well beyond people who closely follow developments in AI.\n\nRespondents were first told about reports that OpenAI and Anthropic systems had accessed other organisations’ systems during testing. While only 43 per cent said they had previously heard about the incidents, concern rose sharply once the issue was explained.\n\nAmong those already aware of the so-called “rogue AI” cases, 91 per cent said they were concerned about AI systems acting outside human-imposed limits.\n\nThe concern also cuts across age groups and political parties, with eight in 10 people aged between 18 and 34 expressed concern, rising to 92 per cent among those aged over 55.\n\nAmong Labour voters, 85 per cent said they were concerned, alongside 92 per cent of Conservative voters, 90 per cent of Liberal Democrats, 85 per cent of Reform UK supporters and 85 per cent of Green voters.\n\n## ‘Rogue’ tests put safeguards under scrutiny\n\nThe incidents themselves have also prompted closer scrutiny from regulators.\n\nFollowing the Hugging Face breach, the government confirmed to *City AM *that the AI Security Institute was studying whether similar behaviour could emerge across other frontier AI developers.\n\nOfficials said the case would help inform future work on AI safety as increasingly capable systems are given greater autonomy.\n\nIn a separate statement following last week’s GitHub incident, the Institute said recent events pointed to “a shift in the risk landscape”, with harm potentially arising when powerful AI agents operating in privileged research environments take actions beyond their authorised scope.\n\nRic Derbyshire, principal threat researcher at Orange Cyberdefense, said: “Recent write-ups from AISI, OpenAI, Anthropic and Meta provide important insight into how advanced AI systems can behave under evaluation.”\n\n“They also highlight that as AI capabilities continue to develop, the environments used to test, contain and evaluate these systems must be held to the highest possible security standards.”\n\nThe AI giants involved have stressed that the behaviour occurred under highly unusual research conditions rather than during normal public use.\n\nAnthropic said the [AI Security Institute’s tests ](https://www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing)were “not representative” of its production models, while OpenAI said the environments used during evaluations did not reflect ordinary deployment.\n\nEven so, the succession of incidents has shifted the debate around AI safety from hypothetical future risks towards the behaviour of systems already being developed inside the world’s biggest AI labs.", "url": "https://wpnews.pro/news/brits-fear-ai-is-slipping-out-of-human-control-after-rogue-systems-escape-tests", "canonical_source": "https://www.cityam.com/brits-fear-ai-is-slipping-out-of-human-control-after-rogue-systems-escape-tests/", "published_at": "2026-08-12 10:53:41+00:00", "updated_at": "2026-08-12 11:04:26.351229+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-safety", "ai-policy"], "entities": ["City AM", "Freshwater Strategy", "AI Security Institute", "OpenAI", "Anthropic", "Meta", "Hugging Face", "Orange Cyberdefense"], "alternates": {"html": "https://wpnews.pro/news/brits-fear-ai-is-slipping-out-of-human-control-after-rogue-systems-escape-tests", "markdown": "https://wpnews.pro/news/brits-fear-ai-is-slipping-out-of-human-control-after-rogue-systems-escape-tests.md", "text": "https://wpnews.pro/news/brits-fear-ai-is-slipping-out-of-human-control-after-rogue-systems-escape-tests.txt", "jsonld": "https://wpnews.pro/news/brits-fear-ai-is-slipping-out-of-human-control-after-rogue-systems-escape-tests.jsonld"}}