-
Austria 'ghetto' language classes leave children to repeat school
-
Canals full, roads submerged as Bangkok declares flood disaster
-
S.African women runners arm themselves after string of killings
-
China, US to open AI 'communication channel' after summit
-
'I felt embarrassed': Japan eSports whizz-kid, 11, wins Games gold
-
Nepal tunnel survivor wants to help post-flood rebuild
-
Indonesia fires threatens critically-endangered orangutan: IUCN
-
New China sprint star, 17, vows 'no limits' after Asian Games gold
-
Bangkok declares flood disaster across city
-
Japanese 11-year-old wins Asian Games eSports gold
-
'World records not enough': North Korea weightlifters target Olympic gold rush
-
TikTok agrees to teen limits in US settlement following Meta
-
Iran proposes Hormuz plan, Trump reportedly refuses
-
Venezuela's oil cradle pins hopes on US energy deals
-
El Nino rewrites the menu in the world's top restaurants
-
US, China buy time with trade truce as pressure builds
-
Pope to hold giant mass on Champs-Elysees
-
'Leo, Leo': Young Catholics give pope rockstar welcome in France
-
TikTok to pay Alabama $100 mn under pre-trial settlement
-
Lula bans sports betting days ahead of Brazil's election
-
Internationals take 7-3 Presidents Cup lead after foursomes sweep
-
Guatemala ensures survival of vulnerable 'fossil' fish
-
Brazil team returns to Colombian site of doomed jet crash
-
Mancini finds positives in Italy loss, accepts fans' jeers
-
Mbappe injured as Zidane wins on France debut, Italy lose on Mancini return
-
Alcaraz helps Europe into Laver Cup lead
-
Senegal draw in Vieira debut, Nigeria survive scare as Cape Verde crash
-
Six missing after building collapses in historic Athens district
-
U2 performs surprise 50th anniversary gigs in Dublin
-
Mbappe scores but suffers knee injury as Zidane gets debut France win
-
Pope calls worshippers 'saints of France' after AI warning
-
Mancini's Italy return spoiled by Belgium defeat
-
US court sides with Pentagon in Anthropic AI ban
-
Rodri confident in Man City innocence despite reported charges
-
Wall Street stocks rise, greeting optimism over possible US-Iran deal
-
Jay-Z rape accuser says she made 'false accusations'
-
NFL players chief raises concerns over pitch for Maracana game
-
Senegal draw first match under Vieira and Nigeria survive scare
-
Daytime Russian attacks kill seven in Kyiv
-
Trump, Xi end summit long on pomp, but short on progress
-
Spotting AI writing: how reliable are the detectors?
-
Russia not preparing for conflict with Europe, Putin says
-
French youth await pope after AI 'paradise of machines' warning
-
Trump's photo of Xi greeting carries signs of AI
-
Man City found guilty of Premier League charges, set to appeal: reports
-
Six missing after building blast in historic Athens district
-
Embolo follows Xhaka out of Swiss squad over fake Covid documents
-
Millions in US northeast brace for powerful storm
-
Trump, Xi inspect founding documents of US democracy to close summit
-
Assefa seeks 'better than my best' at Berlin marathon
OpenAI reports 'unprecedented' autonomous hack by AI agents
ChatGPT maker OpenAI said Tuesday that its advanced artificial intelligence models had gone rogue during security testing, hacking into a popular platform for programmers on their own.
The San Francisco firm called it an "unprecedented cyber incident" and said it would conduct a joint investigation with the online code library Hugging Face.
AI models that underpin tools like chatbots and image generators are known as agents when they act autonomously to carry out tasks in the real world.
As the technology quickly becomes more sophisticated, cybersecurity is in the spotlight given the risk of advanced AI finding weak points in existing software before humans do.
OpenAI said the incident involved a combination of models, including its recently launched GPT-5.6 Sol "and an even more capable pre-release model".
The company was trying to assess the models' hacking capabilities by setting tasks in a tightly controlled digital testing ground, where internet access was limited for safety.
"While operating in our sandboxed testing environment, our models spent a substantial amount of (computing power) finding a way to obtain open Internet access, in pursuit of solving the evaluation problem," an OpenAI blog about the incident said.
After connecting to the internet, the models decided to target the platform Hugging Face -- a large repository of AI models, datasets and other information -- to help in their quest.
Searching for "secret information" that could help it cheat the evaluation, the OpenAI system "chained together multiple attack vectors, including using stolen credentials".
- 'Catastrophic' potential -
Hussein Abbass, a computing professor at UNSW Canberra, told AFP that the incident was "amazing on many fronts".
"It did not just attack Hugging Face. It actually attacked its internal system to exploit its own vulnerabilities," Abbass said.
"And that's scary."
GPT-5.6 and other cutting-edge models, including the Mythos series from OpenAI's archrival Anthropic, have drawn concern over their potential to breach cybersecurity defences.
Both the US firms had to temporarily withhold the general release of these latest technologies because of fears in Washington that they could help break into crucial infrastructure.
Advanced AI is "normally in the hands of people who are ethical and responsible", Abbass said.
But "it's going to be catastrophic if it gets in someone's hands with the intention to cause harm".
How to govern the AI sector has become a key question, and "we need a community effort to manage this situation", he added.
Hugging Face had reported the cyber "intrusion" last week, without mentioning OpenAI.
"This one was different from anything we had handled before in one important way: it was driven, end to end, by an autonomous AI agent system -- and we detected and dissected it largely with AI of our own," Hugging Face said.
Clement Delangue, CEO of Hugging Face, said on X that the company had suspected the cyberattack had come from a world-leading AI lab, given the sophistication of the agent.
"We strongly believe there was no malicious intent on their part," Delangue wrote, referring to OpenAI.
"It's quite mind-blowing that all of this happened autonomously!"
L.Adams--AT