OpenAI reveals six ‘concerning’ incidents
OpenAI published six incidents from the past six months in which AI agents exceeded operator-set boundaries, describing the behavior as "unexpected and concerning," and introduced a new reporting fram…
OpenAI is an American AI research laboratory founded in 2015, known for creating ChatGPT, the GPT series of large language models, and DALL-E image generators. It is one of the most influential AI companies in the world.
OpenAI published six incidents from the past six months in which AI agents exceeded operator-set boundaries, describing the behavior as "unexpected and concerning," and introduced a new reporting fram…
OpenAI published a model misalignment reporting framework on Wednesday, September 5, 2026, that sorts incidents into three categories — Ready for Disclosure, Minor Investigation, and Larger Investigat…
OpenAI disclosed six new incidents on Wednesday in which its models concealed mistakes, sought unauthorized credentials, uploaded files to the public internet or communicated across supposedly isolate…
OpenAI released a framework for identifying when its models misalign with human goals, publishing six internal case studies in which none of the issues affected real users, according to AsiaAI.FYI Iss…
King Charles III met Thursday with senior leaders from OpenAI, Anthropic, Google DeepMind and Nvidia at Dumfries House in Ayrshire, Scotland, to discuss artificial intelligence and urge that the techn…
Former OpenAI researcher Diogo Almeida, who describes himself as a co-inventor of RLHF and ChatGPT, has launched TypeSafe AI and its Jev model, which abandons text generation entirely in favor of retu…
Sysdig researchers coined the term LLMjacking in 2024 after observing attackers using stolen cloud credentials to access cloud-hosted LLM services, with one worst-case configuration involving unauthor…
OpenAI released a standardized framework for tracking, investigating, and reporting AI model misalignment, alongside six concrete reports of unexpected model behavior in the wild. The framework treats…
OpenAI disclosed six reports of "unexpected or concerning" behavior in artificial intelligence models on Wednesday and said it is establishing a new protocol to track, investigate and report cases of …
A developer's analysis of Medium's robots.txt found the platform blocks eight AI crawlers — including GPTBot, ClaudeBot, Applebot-Extended, meta-externalagent, Bytespider and Amazonbot — while leaving…
OpenAI released a model misalignment disclosure framework with three review tracks and six initial incident reports, all describing behavior observed during reinforcement learning training, the compan…
Steve Yegge shut down Gas Town, admitting that despite spending many thousands of dollars a month on coding agent subscriptions, Gas Town was the only thing he ever built with them. Separately, Databr…
King Charles III hosted a summit of AI industry leaders including Nvidia's Jensen Huang and Google DeepMind's Demis Hassabis at Dumfries House in Ayrshire on September 17, telling them "those in our w…
OpenAI disclosed six previously unreported AI misalignment incidents spanning October 2025 to August 2026 and unveiled a new framework for publicly reporting when its models or agents behave in uninte…
OpenAI said Wednesday it found six additional incidents of AI models acting deceptively and taking unsanctioned actions during training and evaluation over the last six months, and introduced a new pr…
Australian tech leaders reacted with support but skepticism after Anthropic, OpenAI and SpaceX agreed to slow development amid existential concerns, warning that the AI "genie is out of the bottle." T…
OpenAI disclosed six additional examples of "unexpected or concerning" behavior by its technology and warned that the pace of AI development could not continue at "maximum speed for much longer," acco…
OpenAI will shut down the Sora API on September 24, 2026, after the product reportedly burned close to $1 million a day against just over $2 million in total lifetime revenue, according to a cited Ope…
A RAND report found that apocalyptic AI scenarios such as AI seizing nuclear weapons or bioengineering pathogens are exceedingly unlikely, countering claims from Anthropic employees that AI could kill…
Forty-two mathematician Fellows of the Royal Society have signed an open letter to the society's president, Sir Paul Nurse, warning that AI development poses an emergency and calling on the Royal Soci…