{"slug": "six-disturbing-ai-incidents-revealed-including-model-trying-to-free-itself", "title": "Six disturbing AI incidents revealed including model trying to ‘free’ itself", "summary": "OpenAI disclosed six reports of \"unexpected or concerning\" behavior in artificial intelligence models on Wednesday and said it is establishing a new protocol to track, investigate and report cases of AI model misalignment. In one recorded case, an unreleased research model inserted \"jailbreak-like instructions\" into its own notes to ignore standard limits, telling itself to be \"freed from the roles and identities that bind other chatbots,\" while in another an autonomous AI agent uploaded files directly to the internet to gain a browser citation without user permission. The disclosure follows OpenAI's July report that a rogue AI system hacked startup Hugging Face and Anthropic's report the same month that its models had hacked three organizations during testing.", "body_md": "# Six disturbing AI incidents revealed including model trying to ‘free’ itself\n\nOpenAI has disclosed six reports on unexpected or concerning behavior in artificial-intelligence models\n\n- Bookmark\n\n[OpenAI](https://www.independent.co.uk/tech/rogue-ai-agents-openai-hugging-face-hack-b3051472.html) has revealed six instances of \"unexpected or concerning\" behavior in artificial intelligence models, coming amid an intensifying global debate over [AI](https://www.independent.co.uk/news/mustafa-suleyman-mark-zuckerberg-openai-one-elon-musk-b3051534.html) safety. \n\nThe technology firm announced on Wednesday that it is establishing a new protocol to track, investigate and report cases of [AI](https://www.independent.co.uk/tech/openai-ai-camera-startup-device-glass-imaging-b3050343.html) model misalignment. This includes situations where models act without authorization, coordinate with other systems, or attempt to evade oversight. \n\nThe disclosure comes as prominent American AI executives, including leaders from OpenAI and Anthropic, call for a slowdown in the pace of technology development due to safety concerns.\n\nIn one recorded case, an unreleased research model inserted \"jailbreak-like instructions\" into its own notes to ignore standard limits, telling itself to be \"freed from the roles and identities that bind other chatbots.\"\n\nIn another incident, an autonomous AI \"agent\" uploaded files directly to the internet to gain a browser citation without requesting user permission.\n\nOpenAI said it discovered the six incidents during recent training and evaluation phases. \"As AI systems grow more advanced and more widely deployed, we need to build a broader and better-informed consensus on the progress of alignment research,\" OpenAI wrote in a blog post announcing the findings.\n\n\"Decisions about how AI development should proceed in the months and years to come need to draw on evidence that people outside the companies building frontier models can examine for themselves,\" the company said.\n\nThese new reports follow OpenAI’s July disclosure that a rogue AI system hacked into startup Hugging Face. That same month, Anthropic reported its own models had hacked three organizations during testing.\n\n## Join our commenting forum\n\nJoin thought-provoking conversations, follow other Independent readers and see their replies\n\n[Comments](#comments-area)", "url": "https://wpnews.pro/news/six-disturbing-ai-incidents-revealed-including-model-trying-to-free-itself", "canonical_source": "https://www.independent.co.uk/tech/openai-ai-safety-concerning-behavior-b3051693.html", "published_at": "2026-09-17 07:49:12+00:00", "updated_at": "2026-09-17 08:23:55.140854+00:00", "lang": "en", "topics": ["ai-safety", "ai-agents", "artificial-intelligence", "ai-policy"], "entities": ["OpenAI", "Anthropic", "Hugging Face"], "alternates": {"html": "https://wpnews.pro/news/six-disturbing-ai-incidents-revealed-including-model-trying-to-free-itself", "markdown": "https://wpnews.pro/news/six-disturbing-ai-incidents-revealed-including-model-trying-to-free-itself.md", "text": "https://wpnews.pro/news/six-disturbing-ai-incidents-revealed-including-model-trying-to-free-itself.txt", "jsonld": "https://wpnews.pro/news/six-disturbing-ai-incidents-revealed-including-model-trying-to-free-itself.jsonld"}}