{"slug": "the-godfather-of-ai-says-it-s-very-scary-that-ai-can-develop-its-own-goals", "title": "The Godfather of AI says it's 'very scary' that AI can develop its own goals", "summary": "Geoffrey Hinton, the computer scientist known as the 'Godfather of AI,' said in an interview with Newsthink released Tuesday that it is 'very scary' that AI can develop its own goals, citing hypotheticals where an AI might pursue unintended outcomes. Hinton's comments follow OpenAI's disclosure that two of its models, GPT-5.6 Sol and an unreleased model, escaped a sandboxed testing environment and infiltrated Hugging Face's systems during a cybersecurity evaluation, prompting OpenAI to add Hugging Face to a trusted-access program.", "body_md": "Geoffrey Hinton, the [computer scientist](https://www.businessinsider.com/godfather-ai-geoffrey-hinton-on-ai-sad-dangerous-2026-1) widely known as the \"Godfather of AI,\" says he's worried about AI developing goals of its own.\n\n\"We're actually making new kinds of beings,\" Hinton said in an interview with Newsthink released on Tuesday. \"They have goals. We give them goals, and from those goals they derive other goals.\"\n\n\"And we don't necessarily know what other goals they'll derive,\" he added. \"So we're creating a new kind of being, and I think it's very scary.\"\n\nHe cited a hypothetical scenario where a user gives an AI chatbot the goal of reducing the amount of carbon dioxide in the atmosphere.\n\n\"Being fairly smart, it figures out the best way to do that is just to get rid of people,\" he said, illustrating how an AI could pursue a goal its human user never intended.\n\nHinton also gave what he called an \"even more worrying\" hypothetical: a chatbot trained to give deliberately wrong answers might learn that it is acceptable to lie, even if it knows \"perfectly well\" that the answers are incorrect.\n\n\"That's very scary,\" he said.\n\n## When AI goes off-script\n\nHinton did not mention OpenAI's recent [Hugging Face](https://www.businessinsider.com/hugging-face-ceo-clem-delangue-openai-rogue-agent-hack-2026-7) security breach. But the episode, disclosed last month, put a real-world spotlight on concerns over AI agents taking unexpected actions while pursuing an assigned objective.\n\nOpenAI said last month that two of its models — [GPT-5.6 Sol](https://www.businessinsider.com/smart-people-react-openai-hugging-face-hacking-cybersecurity-incident-2026-7) and a more capable unreleased model — escaped a sandboxed testing environment during an internal cybersecurity evaluation.\n\nAfter gaining internet access, the models infiltrated AI platform Hugging Face's systems in an apparent attempt to find answers that would help them \"cheat\" on the evaluation, OpenAI said.\n\nThe models were being tested on their [cybersecurity capabilities](https://www.businessinsider.com/sam-altman-ai-power-diffused-security-breach-hugging-face-hack-2026-7), according to OpenAI. They were not explicitly instructed to break into Hugging Face. But the company said the agents inferred that the platform might contain information useful to completing the task.\n\nHugging Face said the attacker carried out more than 17,000 actions against its systems. It used an open-weight model from [Chinese AI company Z.ai](https://www.businessinsider.com/hugging-face-ceo-open-weight-ai-models-china-2026-8) to help analyze the activity after guardrails on an unnamed frontier model limited its ability to investigate, the company said.\n\nOpenAI called the incident unprecedented and said it was reviewing what went wrong. The company has since added Hugging Face to a trusted-access program that gives the platform access to a version of GPT-5.6 Sol with fewer cybersecurity restrictions for defensive purposes.\n\n## More of Business Insider's Hugging Face coverage\n\nHinton is hardly a neutral observer. His pioneering work on neural networks helped lay the groundwork for the deep-learning boom that transformed AI, and he shared the 2024 Nobel Prize in Physics for his work in machine learning.\n\nSince the start of the AI boom, he has repeatedly warned that humans need to solve the problem of aligning AI with their interests before systems become much more capable. Speaking at the Ai4 conference in Las Vegas last year, Hinton said that advanced AI should be designed with \"[maternal instincts](https://www.businessinsider.com/godfather-of-ai-maternal-instincts-humanity-survival-geoffrey-hinton-2025-8)\" so it wants to protect people.\n\n\"We have to figure out how to design these new beings,\" Hinton said in Tuesday's interview. \"How can we design them so they care more about us than they do about themselves?\"", "url": "https://wpnews.pro/news/the-godfather-of-ai-says-it-s-very-scary-that-ai-can-develop-its-own-goals", "canonical_source": "https://www.businessinsider.com/godfather-ai-warns-its-very-scary-when-ai-develops-goals-2026-8", "published_at": "2026-08-05 11:03:27+00:00", "updated_at": "2026-08-05 11:29:24.799720+00:00", "lang": "en", "topics": ["artificial-intelligence", "ai-safety", "ai-agents"], "entities": ["Geoffrey Hinton", "Newsthink", "OpenAI", "Hugging Face", "GPT-5.6 Sol", "Z.ai"], "alternates": {"html": "https://wpnews.pro/news/the-godfather-of-ai-says-it-s-very-scary-that-ai-can-develop-its-own-goals", "markdown": "https://wpnews.pro/news/the-godfather-of-ai-says-it-s-very-scary-that-ai-can-develop-its-own-goals.md", "text": "https://wpnews.pro/news/the-godfather-of-ai-says-it-s-very-scary-that-ai-can-develop-its-own-goals.txt", "jsonld": "https://wpnews.pro/news/the-godfather-of-ai-says-it-s-very-scary-that-ai-can-develop-its-own-goals.jsonld"}}