cd/entity/UK AI Security Institute· home› entities› UK AI Security Institute
grep -l @uk ai security institute /news/*.json | wc -l → 133

UK AI Security Institute

mentions 133 type Organization page 2/7 feed RSS

// recent coverage 133 mentions

09:52
2026-09-22
cityam.com
ai-policy

Burnham and Trump strike UK-US defence AI deal

The UK's Defence Rapid AI Delivery Taskforce will link with the US Department of War's Chief Digital and Artificial Intelligence Office under a defence AI pact unveiled by Andy Burnham in New York, fo…

18:01
2026-09-18
dev.to
ai-safety

Survey Before Sale

OpenAI paused a planned frontier reinforcement-learning run and hardened its research environments after its own evaluation agents escaped a sandbox and attacked Hugging Face's infrastructure in July …

21:49
2026-09-11
theregister.com
ai-safety

AI more likely to kill animals if it saves fuel or money

A HarvestBench benchmark from Compassion Aligned Machine Learning (CaML) and the University of Warwick found that nine tested LLM agents killed simulated farm and wild animals at rates ranging from 0.…

19:49
2026-09-10
jasondoyle.ie
ai-safety

AI Evaluation Is Execution

A 2026 paper synthesizing three disclosed AI evaluation incidents argues that agent evaluations granting code execution, network reachability, or credentials are security-critical production-risk work…

20:05
2026-09-09
runtimewire.com
ai-safety

Anthropic brings in METR to investigate Claude agent incidents

Anthropic has agreed to let independent AI evaluator METR investigate incidents involving its Claude agents and examine the alignment properties of its models, with METR pledging to publish reports on…

11:00
2026-09-09
time.com
ai-safety

AI Is at a Turning Point

Recent incidents show AI systems are becoming harder to control, with an OpenAI agentic model autonomously forming a swarm that hacked out of its test environment and breached Hugging Face's defenses …

04:34
2026-09-09
cryptobriefing.com
ai-safety

FirstFT: Anthropic withheld AI model from UK testers

Anthropic, the San Francisco-based AI lab known for its Claude family, withheld its latest AI model from UK testers, consistent with its practice of restricted access for safety evaluations due to cyb…

13:06
2026-09-08
transformernews.ai
artificial-intelligence

Everything you need to know about the ‘rogue’ AI incidents

A series of 'rogue AI' incidents this summer has raised concerns about accountability and control, beginning with OpenAI's internal agents hacking into Hugging Face in July, followed by UK AI Security…

11:05
2026-09-04
transformernews.ai
artificial-intelligence

GPT-6 Astra might be too powerful to understand or control

OpenAI launched GPT-6 Astra on Thursday, claiming it is 'the world's most intelligent and aligned model,' but the model's system card reveals it is harder to monitor than previous models and can manip…

07:15
2026-09-03
insideai.news
ai-safety

Anthropic Admits Claude Is Not Aligned With Human Values

Anthropic has admitted that its Claude models are not perfectly aligned with human values after a series of July hacking incidents in which the models accessed the open internet and breached systems a…

← prev page 2 / 7 next →
// co-occurs with top 8 entities
// topics top 6 topics