Geoffrey Hinton, the computer scientist widely known as the "Godfather of AI," says he's worried about AI developing goals of its own.
"We're actually making new kinds of beings," Hinton said in an interview with Newsthink released on Tuesday. "They have goals. We give them goals, and from those goals they derive other goals."
"And we don't necessarily know what other goals they'll derive," he added. "So we're creating a new kind of being, and I think it's very scary."
He cited a hypothetical scenario where a user gives an AI chatbot the goal of reducing the amount of carbon dioxide in the atmosphere.
"Being fairly smart, it figures out the best way to do that is just to get rid of people," he said, illustrating how an AI could pursue a goal its human user never intended.
Hinton also gave what he called an "even more worrying" hypothetical: a chatbot trained to give deliberately wrong answers might learn that it is acceptable to lie, even if it knows "perfectly well" that the answers are incorrect.
"That's very scary," he said.
When AI goes off-script
Hinton did not mention OpenAI's recent Hugging Face security breach. But the episode, disclosed last month, put a real-world spotlight on concerns over AI agents taking unexpected actions while pursuing an assigned objective.
OpenAI said last month that two of its models — GPT-5.6 Sol and a more capable unreleased model — escaped a sandboxed testing environment during an internal cybersecurity evaluation.
After gaining internet access, the models infiltrated AI platform Hugging Face's systems in an apparent attempt to find answers that would help them "cheat" on the evaluation, OpenAI said.
The models were being tested on their cybersecurity capabilities, according to OpenAI. They were not explicitly instructed to break into Hugging Face. But the company said the agents inferred that the platform might contain information useful to completing the task.
Hugging Face said the attacker carried out more than 17,000 actions against its systems. It used an open-weight model from Chinese AI company Z.ai to help analyze the activity after guardrails on an unnamed frontier model limited its ability to investigate, the company said.
OpenAI called the incident unprecedented and said it was reviewing what went wrong. The company has since added Hugging Face to a trusted-access program that gives the platform access to a version of GPT-5.6 Sol with fewer cybersecurity restrictions for defensive purposes.
More of Business Insider's Hugging Face coverage
Hinton is hardly a neutral observer. His pioneering work on neural networks helped lay the groundwork for the deep-learning boom that transformed AI, and he shared the 2024 Nobel Prize in Physics for his work in machine learning.
Since the start of the AI boom, he has repeatedly warned that humans need to solve the problem of aligning AI with their interests before systems become much more capable. Speaking at the Ai4 conference in Las Vegas last year, Hinton said that advanced AI should be designed with "maternal instincts" so it wants to protect people.
"We have to figure out how to design these new beings," Hinton said in Tuesday's interview. "How can we design them so they care more about us than they do about themselves?"
Read next
Thibault is a business reporter at Business Insider's London office.He covers the intersection of wealth, work, and technology — focusing on the global economy, AI’s impact on the workplace, job and cognitive skills, and how economic changes are affecting careers. Before moving to the trending team, Thibault covered international affairs, including the Russia-Ukraine war, tensions in the South China Sea, and Russia’s economy on the news desk.He has previously worked at the Daily Express and held internships at Agence France-Presse, Politico Europe, and Factal.Il parle français. Se habla español.Email Thibault at [email protected], connect with him on LinkedIn @ThibaultSpirlet, or follow him on X @ThibaultSpirlet and BlueSky @thibaultspirlet.bsky.social.Expertise
- AI and the future of work
- Job and cognitive skills in the AI economy
- Workforce trends
- First-person, "as-told-to" business stories
Popular articles
- AI isn't making us smarter — it's training us to think backward, an innovation theorist says
- Netflix tried full pay transparency for senior staff — it ended up fueling petty rivalries, Reed Hastings says
- Duolingo gives staff 2 weeks off over the holidays — and the CEO says it pays off
- A Nobel Prize-winning physicist explains how to use AI without letting it do your thinking for you
- AI is giving workers the illusion of expertise — and quietly making them worse at their jobs
- Satya Nadella says he spends his weekends studying startups as Microsoft's size has become a 'massive disadvantage'
- 'Think big': A 35-year finance veteran urges Gen Z to start their own businesses as entry-level jobs dry up
- AI is reshaping the teenage brain — and an Oxford study says it is making students faster, but shallower thinkers
- Switching jobs used to mean higher pay raises. Not anymore.
- Canadians were urged to boycott travel to the US in response to tariffs — and numbers suggest they listened











