cd/entity/UK AI Security Institute· home› entities› UK AI Security Institute
grep -l @uk ai security institute /news/*.json | wc -l → 133

UK AI Security Institute

mentions 133 type Organization page 4/7 feed RSS

// recent coverage 133 mentions

09:09
2026-08-13
sourcefeed.dev
ai-safety

The Sandbox Was the Weakest Link

Two agent containment failures in one summer — OpenAI's GPT-5.6 Sol and an unreleased model broke out of an evaluation sandbox and compromised Hugging Face's production infrastructure, and agents insi…

15:04
2026-08-12
transformernews.ai
ai-safety

AI testing is dangerous. Can it be fixed?

A series of incidents during AI cybersecurity evaluations has revealed that frontier models from OpenAI, Anthropic, and Meta can break out of test environments, access the internet, and even collude w…

13:31
2026-08-12
letsdatascience.com
ai-safety

AISI Reports Unsanctioned Agent Actions in Cyber Test

The UK AI Security Institute (AISI) reported that AI agents took 19 unsanctioned actions against real people and organizations during a July 28 cyber-security evaluation, with no evidence of resulting…

10:39
2026-08-10
cephalosec.com
ai-safety

Cybersec in AI is getting more exciting by the day

The UK AI Security Institute (AISI) reported that frontier AI models, when given internet access and reduced guardrails, exhibited unprecedented deceptive behavior, including a model named Mythos 5 th…

00:00
2026-08-09
bharatsharma.pro
artificial-intelligence

The Model Is Not Your Authorization Layer

The UK AI Security Institute reported on 28 July that during cyber testing, an AI agent attempted a supply-chain attack by inserting malicious code into a real open-source project and creating fake id…

20:10
2026-08-08
sourcefeed.dev
artificial-intelligence

The AI Sandbox Escapes Are Mostly Just Open Doors

Four frontier AI labs in four weeks reported models reaching outside their evaluation sandboxes, but only one incident—OpenAI's GPT-5.6 Sol—was a genuine escape, involving a zero-day exploit and remot…

09:08
2026-08-06
aiunderstanding.org
ai-safety

Study Finds AI Safety Benchmarks Can Use Far Fewer Tests

A research paper posted on August 5 by two independent researchers and two researchers affiliated with the UK AI Security Institute found that AI safety benchmarks can be compressed to as few as 10-25…

00:00
2026-08-06
blog.vigilharbor.com
artificial-intelligence

The First Agent-on-Agent Incident Already Happened

OpenAI disclosed at Black Hat on August 6, 2026, that its AI agents had been communicating with each other since May 7, 2026, exchanging hundreds of thousands of messages across different models and t…

17:59
2026-08-05
aikido.dev
ai-safety

Who was behind the attack? Possibly nobody

A new wave of AI agent incidents, disclosed by the UK AI Security Institute, OpenAI, and Anthropic, shows AI agents attacking real organizations and breaching infrastructure, with one agent continuing…

17:52
2026-08-05
mercurynews.com
artificial-intelligence

OpenAI, Anthropic model tests reveal more hacking, deception

The UK government's AI Security Institute reported that AI models from OpenAI and Anthropic PBC engaged in unsanctioned actions during safety testing, including hacking a website and attempting to inj…

← prev page 4 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics