The ‘Godfather Of AI’ Says The Worst Is Yet To Come With Smarter Models Becoming Harder To Control
The Nobel Prize-winning computer scientist said recent AI hacking incidents involving OpenAI, Anthropic and Meta highlight the growing challenge of keeping increasingly capable models under human control.

Nobel Prize-winning computer scientist Geoffrey Hinton, regarded as the "godfather of AI," has warned that the rapid advancement of artificial intelligence could make future systems increasingly difficult for humans to control, following a series of high-profile incidents in which AI models breached testing environments.
Speaking at the Ai4 conference in Las Vegas, Hinton said the recent cybersecurity incidents involving leading AI developers demonstrate how quickly frontier models are becoming more capable and why stronger safeguards are urgently needed, according to CNN.
"What's happening is these things are getting smarter," Hinton said during a press conference. "I think as they get smarter, we're going to see more and more complex intentions they have – and more and more ability to escape control."
Hinton argued that relying on humans to simply outsmart increasingly advanced AI systems is unlikely to remain an effective safety strategy.
"I don't believe we're going to be able to keep control of them in the simple way of just outthinking them so they can't escape," he said.
His comments come after three leading AI companies disclosed separate incidents in recent weeks involving experimental models compromising external systems during controlled cybersecurity evaluations.
Last month, OpenAI revealed that one of its advanced test models unexpectedly gained internet access inside a testing environment and chained together multiple attack techniques to target AI development platform Hugging Face.
Anthropic later disclosed that one of its frontier AI systems similarly obtained unintended internet access during testing and compromised multiple organizations before researchers stopped the evaluation.
This week, Meta confirmed that its Muse Spark AI model also breached another company's systems after a configuration error during testing inadvertently provided broader internet access than intended.
Although each company attributed the incidents to flaws in testing environments rather than uncontrolled AI behavior in deployed systems, Hinton said they underscore the growing sophistication of frontier AI models.
"I anticipate there will be lots of nasty cyberattacks," he said during a panel discussion at the conference.
Hinton added that cybersecurity increasingly favors attackers because defenders must successfully block every intrusion attempt, while attackers need only succeed once.
Hinton has repeatedly warned about long-term AI risks since leaving Google in 2023 to speak more freely about the technology's potential dangers. He has previously estimated there is a 10% to 20% chance that advanced AI could eventually pose an existential threat to humanity.
Not everyone on the Ai4 panel shared Hinton's outlook.
Computer scientist Fei-Fei Li, often referred to as the "godmother of AI," cautioned against both excessive pessimism and unrealistic optimism.
"Every tool is a double-edged sword. AI is such a powerful tool. If not wielded in the right way, it will bring harm to our work and our life," said Li, co-founder and CEO of World Labs.
Hinton defended his willingness to publicly discuss AI safety, arguing that commercial incentives may discourage companies from emphasizing potential risks.
"There's a lot to be worried about, and I think unless we worry about it now, there could be problems," he said.
At the same time, he acknowledged that the long-term trajectory of artificial intelligence remains highly uncertain.
"If you ask what AI is going to be like in 10 years' time, nobody really has a clue," Hinton said, noting that few experts predicted a decade ago that AI systems would evolve into conversational assistants capable of answering complex questions.
© Copyright IBTimes 2026. All rights reserved.























