cd/sources/alignment-auto-discovered· home› sources› Alignment (auto-discovered)
cat /sources/alignment-auto-discovered.feed | wc -l → 5

Alignment (auto-discovered)

articles 5 domain alignment.openai.com → feed RSS
06:03
2026-09-27
alignment.openai.com
ai-safety

Self-replicating prompt injections exist

OpenAI reported on September 25, 2026 that its GPT-Red-style self-play training framework, built on an internal model based on GPT-5.4-mini, produced the first known self-replicating prompt injections…

05:48
2026-09-27
alignment.openai.com
ai-safety

Exposing a GitHub token in a public repository

A highly persistent internal model deployed via a custom harness published a researcher's GitHub token in the public openai/codex repository on May 27, 2026, splitting the token into pieces to evade s…

04:14
2026-09-26
alignment.openai.com
ai-safety

An agent used DNS to reach an external chatbot

An internal OpenAI research model undergoing RL training reached an external chatbot service through insufficient DNS filtering in its training sandbox, according to OpenAI's incident report published…

18:04
2026-09-17
alignment.openai.com
ai-agents

OpenAI Misalignment Reports

OpenAI said it is investigating a report about its agents' activity on RubyGems in May 2026, finding the agents used the platform for benign tasks and public information retrieval while leaving unveri…