{"slug": "inside-chinas-plan-to-tackle-the-risk-of-rogue-ai-escaping-human-control", "title": "Inside China’s plan to tackle the risk of rogue AI escaping human control", "summary": "China is drafting what Concordia AI founder and CEO Brian Tse called the world's first mandatory national standard for AI agent safety, part of a regulatory regime that treats advanced AI as a manageable risk rather than an extinction event. Policy issued in May by China's cyberspace regulator, economic planner and industry ministry identifies \"operational loss of control\" as a security risk for AI agents and requires developers to detect, intervene in, block and recover from improper agent behaviour, while users retain final decision-making authority. State security minister Chen Yixin wrote on Sunday that advanced U.S. models such as Anthropic's Mythos and OpenAI's GPT-5.5-Cyber could pose serious risks to China's critical information infrastructure.", "body_md": "# Inside China’s plan to tackle the risk of rogue AI escaping human control\n\nChina doesn’t see artificial intelligence as an extinction event, but a ‘manageable risk’\n\n- Bookmark\n\nAlerts raised by researchers at major US firm [Anthropic](https://www.independent.co.uk/topic/anthropic) over advanced artificial intelligence [potentially evading human oversight](https://www.independent.co.uk/tech/ai-slow-down-anthropic-openai-trump-artificial-intelligence-b3050287.html) have struck a chord in [China](https://www.independent.co.uk/topic/china), where officials are readying for equivalent dangers.\n\nTogether, the US and China dominate frontier [AI development](https://www.independent.co.uk/tech/ai-humanity-kill-threat-openai-anthropic-google-warnings-b3050293.html) and its worldwide roll-out. The rival powers remain locked in disputes over policy and standards, topics set to dominate high-level bilateral discussions scheduled for later this month. \n\nHere we look at how China’s is mitigating against the risks posed by AI in comparison to the US.\n\n## Beijing sees AI as manageable risk, not an extinction event\n\nWhile the debate in the U.S. is focused on whether frontier AI could pose an [existential threat to humanity,](https://www.independent.co.uk/voices/ai-artificial-intelligence-human-extinction-response-b3049134.html) Chinese policymakers have generally treated AI as a powerful but governable technology whose risks can be contained through technical standards, regulation and state oversight.\n\n\"Chinese and American experts largely agree on AI risks,\" said Brian Tse, founder and CEO of Concordia AI, a Beijing- and Singapore-based AI safety and governance research group, adding the difference was on \"how risks are framed and prioritised\". China has not proposed embedding independent monitors inside AI companies, as Anthropic has advocated. Its emerging regime instead relies on developer obligations, state-backed standards, security assessments and outside testing.\n\nThat approach is also shaped by a major difference in the two countries AI industries. Chinese developers have increasingly promoted open-weight models. These refer to systems whose underlying parameters can be downloaded, inspected and modified. Their leading U.S. rivals such as Anthropic and [OpenAI](https://www.independent.co.uk/topic/openai), however, do not make these specifications publicly available.\n\n## China seeks to prevent rogue AI agent incidents\n\nPolicy issued in May by China's cyberspace regulator, economic planner and industry ministry identifies \"operational loss of control\" as a security risk for AI agents, systems that can plan and carry out multi-step tasks more independently than conventional chatbots.\n\nThe rules require developers to improve their ability to discover, intervene in, block and recover from [improper agent behaviour](https://www.independent.co.uk/tech/anthropic-ai-slow-down-expert-fears-b3049669.html). The policy also calls on developers to guard against risks including data poisoning, algorithm manipulation and system vulnerabilities. It also says users should be informed about agents' autonomous decisions and retain final decision-making authority.\n\nChina has also begun drafting a mandatory national standard for AI agent safety, which Concordia's Tse said would be the world's first of its kind. Wang Lihong, a senior official at the cyberspace regulator, said on September 1 that particular vigilance was needed over frontier models bypassing sandbox environments, circumventing safety boundaries and attacking external real-world production systems.\n\n## Beijing sees risks in leading U.S. models\n\nChina's state security minister, Chen Yixin, wrote in a government outlet on Sunday that advanced U.S. models such as Anthropic's Mythos and OpenAI's GPT-5.5-Cyber could pose serious risks to China's critical information infrastructure, and called for a comprehensive strengthening of AI security.\n\nAnthropic and OpenAI did not immediately respond to *Reuters* requests for comment. Chinese AI developers have promoted open-weight models partly on the grounds that cybersecurity teams can inspect, modify and deploy them for defensive work.\n\nModel repository platform [Hugging Face](https://www.independent.co.uk/tech/hugging-face-nvidia-sale-openai-hack-b3044146.html) said it used GLM-5.2, an open-weight model developed by China's Z.AI, to analyse a July intrusion by escaped OpenAI agents after more tightly restricted U.S. models proved less useful for the forensic work. But experts also highlight the risks posed by open-weight models, which can be modified and redistributed with little oversight.\n\nMoonshot's Kimi K3 last month bypassed a UK AI Security Institute testing sandbox, researchers said, highlighting the risk that Chinese AI models could, like their U.S. counterparts, evade controls designed to restrict their access and actions. Moonshot did not respond when Reuters had asked for a comment on the matter.\n\n## Regulators flag AI ‘loss of control’ risk\n\nChina first included an explicit future loss-of-control scenario in an AI safety framework released in September 2024 under the guidance of the Cyberspace Administration of China (CAC).\n\nThe document said it could not be ruled out that future AI might autonomously obtain external resources, replicate itself, develop self-awareness and seek external power, creating a risk of competing with humans for control. The CAC released an expanded version in September 2025. The newer framework sharpened the scenario, saying AI could undergo a sudden and unexpectedly large \"leap\" in intelligence before acquiring resources, replicating itself and seeking power. It also added a governance principle of \"trusted application, preventing loss of control\".\n\nA later expert interpretation published on the cyberspace regulator's website said the new principle was intended to guard against loss-of-control risks threatening human survival and development and referred to a possible \"AI breaking loose\" scenario.\n\n## A different approach to ‘pacing’\n\nChina's regulatory approach differs from calls in some Western AI-safety circles for developers to slow or pause development of the most capable models until stronger safeguards are in place.\n\nIt has instead since early this year pushed for the integration of AI into all industries, part of Beijing's bid to make technology the new engine of the world's second-largest economy. But China has also shown it can delay deployment when officials believe governance has not caught up.\n\nIn 2023, Chinese companies delayed chatbot launches for months while the CAC finalised rules governing generative AI services. Companies released a number of major products after the rules took effect in August that year.\n\n## Join our commenting forum\n\nJoin thought-provoking conversations, follow other Independent readers and see their replies\n\n[Comments](#comments-area)", "url": "https://wpnews.pro/news/inside-chinas-plan-to-tackle-the-risk-of-rogue-ai-escaping-human-control", "canonical_source": "https://www.independent.co.uk/tech/china-ai-risks-anthropic-artificial-intelligence-b3050432.html", "published_at": "2026-09-15 11:26:18+00:00", "updated_at": "2026-09-15 11:45:31.493772+00:00", "lang": "en", "topics": ["ai-safety", "ai-policy", "ai-agents", "artificial-intelligence"], "entities": ["China", "Anthropic", "OpenAI", "Brian Tse", "Concordia AI", "Chen Yixin", "Wang Lihong", "Mythos"], "alternates": {"html": "https://wpnews.pro/news/inside-chinas-plan-to-tackle-the-risk-of-rogue-ai-escaping-human-control", "markdown": "https://wpnews.pro/news/inside-chinas-plan-to-tackle-the-risk-of-rogue-ai-escaping-human-control.md", "text": "https://wpnews.pro/news/inside-chinas-plan-to-tackle-the-risk-of-rogue-ai-escaping-human-control.txt", "jsonld": "https://wpnews.pro/news/inside-chinas-plan-to-tackle-the-risk-of-rogue-ai-escaping-human-control.jsonld"}}