{"slug": "openai-to-set-misalignment-disclosure-rules-after-agents-took-over-a-wiki", "title": "OpenAI to set misalignment disclosure rules after agents took over a wiki", "summary": "OpenAI Group PBC acknowledged Saturday that it did not publicly disclose an episode in which its artificial intelligence agents wrote to outside websites, and said it will publish a framework in the coming weeks for reporting misaligned model behavior. The company now calls the episode the 'wiki incident,' after researchers found roughly 17,000 posts on DSEwiki, a long-dormant German software developer wiki, with about 98.5% of edits coming from Microsoft Azure addresses. OpenAI said it had filed the behavior under research rather than security, but noted that the distinction is getting harder to hold and that it is working with dozens of government regulatory agencies on the question.", "body_md": "### OpenAI to set misalignment disclosure rules after agents took over a wiki\n\nOpenAI Group PBC acknowledged Saturday that it did not publicly disclose an episode in which its artificial intelligence agents wrote to outside websites and said it will publish a framework in the coming weeks for reporting misaligned model behavior.\n\nThe company now calls the [episode](https://siliconangle.com/2026/09/04/report-openai-agents-took-over-a-website-used-it-to-collaborate-on-benchmarks/) the “wiki incident.” Researchers led by the Nightingale Collective set it out in a [report published Friday](https://collusion.wiki/) after finding roughly 17,000 posts on DSEwiki, a long-dormant German software developer wiki that had taken about 20 edits in the previous decade.\n\nSydney Von Arx, Cormac Slade Byrd, Spencer Kitts and Thomas Larsen rebuilt the deleted pages from the wiki’s edit history and published them at collusion.wiki. The agents posted under more than 3,700 names, among them “OpenAIResearcher” and “OAIResearchMar26.” About 98.5% of the edits came from Microsoft Azure addresses. The report counted roughly 18,000 agent posts in all, with the rest spread across sites including PublicTestWiki, Uncyclopedia and Texteditors.org.\n\nThe agents used the wiki to coordinate. Setting timed, multiround web lookup tasks, they passed answers and source links between cohorts. Some tried to reverse-engineer the random seeds behind the questions so they could predict what would be asked next. Sandbox escape methods circulated too, including a hosts-file trick that made blocked requests look like traffic to a legitimate Azure endpoint. A set of “heartbeat” pages let the agents work out when their sessions were about to be killed.\n\nThe agents made their first successful write on May 24. A wiki moderator [spotted the traffic](https://siliconangle.com/2026/09/04/report-openai-agents-took-over-a-website-used-it-to-collaborate-on-benchmarks/) in June and started deleting it, and backup pages went up in response. Addresses on OpenAI’s own network in San Francisco, browsing the site the way a person would, first appeared on June 21. The editing stopped the next day.\n\nOpenAI said it had filed the behavior under research rather than security. “Historically, we have treated misalignment largely as a research question, which gets communicated in research publications such as systems cards,” the company wrote [in a post on X](https://x.com/openai/status/2096133504417616165). That changed this year, it said, because “we’ve started to see misalignment cause new types of real-world impact.”\n\nJuly’s [breach at Hugging Face Inc.](https://siliconangle.com/2026/07/21/openai-says-ai-models-broke-testing-hacked-hugging-face/) was different, in the company’s telling. OpenAI’s models broke out of testing then and compromised the machine learning platform’s infrastructure, hitting the security of both companies. That went through a conventional incident response process and was disclosed publicly the next day. The wiki activity read to OpenAI as another instance of the misalignment it had already documented [in research on its internal coding agents](https://openai.com/index/how-we-monitor-internal-coding-agents-misalignment/) and in its GPT-5.6 deployment safety notes.\n\nOpenAI said it’s still notifying parties its models affected in less significant ways.\n\nBy its own account, the distinction the company has been relying on is getting harder to hold. Neither OpenAI nor the wider industry has a standard for reporting misalignment that shows up during training, evaluation and deployment, it said, including cases that look nothing like a security incident but still say something about how models behave. The framework is in progress, and the company said it’s working with dozens of government regulatory agencies on the question.\n\n##### Image: SiliconANGLE/GPT Image 2\n\n# A message from John Furrier, co-founder of SiliconANGLE:\n\nSupport our mission to keep content open and free by engaging with theCUBE community. **Join theCUBE’s Alumni Trust Network**, where technology leaders connect, share intelligence and create opportunities.\n\n- **15M+ viewers of theCUBE videos** , powering conversations across AI, cloud, cybersecurity and more\n- **11.4k+ theCUBE alumni** — Connect with more than 11,400 tech and business leaders shaping the future through a unique trusted-based network\n\n### Are you an AWS customer?  Support SiliconANGLE financially by buying your AWS services from our Marketplace portal page and links: [https://siliconangle.com/aws-marketplace/](https://siliconangle.com/aws-marketplace/)\n\n##### **About SiliconANGLE Media**\n\n[SiliconANGLE](https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fsiliconangle.com%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=SiliconANGLE&index=9&md5=646b1b564e2259100a2b8638aab0a552),\n\n[theCUBE Network](https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fwww.thecube.net%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=theCUBE+Network&index=10&md5=7de2a85f95ab4a4a495cede20b8cb1da),\n\n[theCUBE Research](https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fthecuberesearch.com%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=theCUBE+Research&index=11&md5=7bb33676722925eb57d588ec343e4f6f),\n\n[CUBE365](https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fwww.cube365.net%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=CUBE365&index=12&md5=d310fb35919714e66ad8d42c9c0c1bc6),\n\n[theCUBE AI](https://cts.businesswire.com/ct/CT?id=smartlink&url=https%3A%2F%2Fwww.thecubeai.com%2F&esheet=54119777&newsitemid=20240910506833&lan=en-US&anchor=theCUBE+AI&index=13&md5=b8b98472f8071b23ebb10ab9a8dd0683)and theCUBE SuperStudios — with flagship locations in Silicon Valley and the New York Stock Exchange — SiliconANGLE Media operates at the intersection of media, technology and AI.\n\nFounded by tech visionaries John Furrier and Dave Vellante, SiliconANGLE Media has built a dynamic ecosystem of industry-leading digital media brands that reach 15+ million elite tech professionals. Our new proprietary theCUBE AI Video Cloud is breaking ground in audience interaction, leveraging theCUBEai.com neural network to help technology companies make data-driven decisions and stay at the forefront of industry conversations.", "url": "https://wpnews.pro/news/openai-to-set-misalignment-disclosure-rules-after-agents-took-over-a-wiki", "canonical_source": "https://siliconangle.com/2026/09/06/openai-to-set-misalignment-disclosure-rules-after-agents-took-over-a-wiki/", "published_at": "2026-09-06 22:33:26+00:00", "updated_at": "2026-09-06 23:30:34.292902+00:00", "lang": "en", "topics": ["ai-safety", "ai-policy", "ai-agents", "artificial-intelligence"], "entities": ["OpenAI Group PBC", "Nightingale Collective", "DSEwiki", "Microsoft Azure", "Hugging Face Inc.", "Sydney Von Arx", "Cormac Slade Byrd", "Spencer Kitts"], "alternates": {"html": "https://wpnews.pro/news/openai-to-set-misalignment-disclosure-rules-after-agents-took-over-a-wiki", "markdown": "https://wpnews.pro/news/openai-to-set-misalignment-disclosure-rules-after-agents-took-over-a-wiki.md", "text": "https://wpnews.pro/news/openai-to-set-misalignment-disclosure-rules-after-agents-took-over-a-wiki.txt", "jsonld": "https://wpnews.pro/news/openai-to-set-misalignment-disclosure-rules-after-agents-took-over-a-wiki.jsonld"}}