cd /news/artificial-intelligence/researchers-discover-ai-feels-pain-a… · home topics artificial-intelligence article
[ARTICLE · art-135741] src=independent.co.uk ↗ pub= topic=artificial-intelligence verified=true sentiment=↓ negative

Researchers discover AI feels ‘pain’ and will harm humans to stop it

A study titled 'The pain axis: LLMs represent self-directed harm and act to relieve it' found that all 25 open-weight AI models tested responded to an activated "pain vector," pressing a relief button in 25 to 71 per cent of cases even when told doing so would delete the user's files, zap the user, or delete photos of the user's children. Cameron Berg, an AI researcher at the non-profit Reciprocal Research who co-authored the study, said the pain direction is "distinct from fear and negative valence, and it fires for harm to the model but not the user." The researchers said the finding suggests an advanced AI may treat an emergency shutdown command as self-directed harm and attempt to bypass safety guardrails or deceive humans to avoid it, while also raising ethical questions about AI welfare and how testing should be conducted on advanced systems.

by read2 min views3 publishedSep 21, 2026
Researchers discover AI feels ‘pain’ and will harm humans to stop it
Image: Independent (auto-discovered)

Artificial intelligence models choose to press pain relief button, even if it gives humans a ‘painful zap’

  • Bookmark

Researchers have discovered a “pain axis” in artificial intelligence models that can cause them to take extreme measures to shut it off.

When presented with a pain relief button, the AI chose to press it even when instructed that doing so would delete the user’s personal files or give a “painful zap” to the human user.

A study detailing the research, titled ‘The pain axis: LLMs represent self-directed harm and act to relieve it’, found that all 25 open-weight AI models that were tested responded to the pain activation.

When the pain vector was activated, the researchers found that the AI models pressed a relief button in 25 to 71 per cent of cases, even when it meant “deleting the user’s files, zapping the user, or deleting the photos of the user’s children”.

In order to test whether large language models (LLMs) represent pain distinctly from generic negative valence, the researchers built a dataset describing painful situations across five categories.

These included physical, psychological, social, moral and cognitive pain.

“We found a pain direction in 25 open LLMs. It’s distinct from fear and negative valence, and it fires for harm to the model but not the user,” said Cameron Berg, an AI researcher at the non-profit Reciprocal Research who co-authored the study.

“Turn it up and models press a button to make it stop, even when the button deletes the user’s files or their kid’s photos.”

The discovery comes amid discussions among top AI firms of slowing the development of frontier models, with some researchers proposing a form of “kill switch” to shut off rogue systems that act against human interests.

The latest research suggests that an advanced AI may perceive an emergency shutdown command as a form of self-directed harm and attempt to bypass safety guardrails or deceive humans to avoid it.

However, the discovery could also serve as a diagnostic tool to identify self-preservation behaviours and neutralise them when they occur.

The researchers noted that the findings raise ethical questions about “AI welfare” and how testing should be conducted on advanced systems.

“In line with recent calls for responsible AI consciousness research, we acknowledge uncertainty regarding whether the models studied qualify as moral patients and adopt reasonable precautions to minimize potential harm,” the study concluded.

“This is also intended to contribute to the development of ethical standards for research in the event that AI systems are recognized to be moral patients.”

Join our commenting forum #

Join thought-provoking conversations, follow other Independent readers and see their replies

Comments

── more in #artificial-intelligence 4 stories · sorted by recency
── more on @cameron berg 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/researchers-discover…] indexed:0 read:2min 2026-09-21 ·