-
Death toll from attack at police HQ in Pakistan rises to 31
-
Drone owners rush to offload devices ahead of Beijing ban
-
Trump says Denmark to give US 'permanent control' over Greenland security
-
Trump's beef policies irk Texas ranchers, testing their loyalty
-
Kennedy Center: future unclear for iconic Washington arts hub
-
Nvidia, OpenAI CEOs to attend Xi dinner at White House
-
Trump claims 'infinite' Greenland security deal with Denmark
-
Anthropic picks Accenture for in-house AI safety evaluation
-
Helicopters battle wildfire threatening Ecuador's capital
-
Balogun returns as Monaco beat Lens to continue strong Ligue 1 start
-
Brentford batter Chelsea to undo Alonso's promising start
-
'Pause' on overt repression in Venezuela, but reforms still needed: HRW
-
Kane hits Bundesliga century as Bayern rout Union
-
Global stocks mixed as yen falls despite Bank of Japan rate hike
-
Trump says CNN, MS NOW, Politico banned from White House
-
Cuba hit with seventh major blackout of the year
-
Hushing and hedging: US companies retreat on climate
-
Murray out with concussion so Wentz will start for Vikings
-
Springboks to face Fiji before 2027 Rugby World Cup
-
UN chief 'deeply regrets' US visa refusal to Abbas for annual meeting
-
Nigeria miners struggled to breathe in cell before 37 died: survivors
-
England quick Carse to face no charges over alleged nightclub assault
-
McIlroy on the charge at PGA Championship as Reed withdraws
-
Mancini brings nine newcomers into the first squad of his Italy comeback
-
McIlroy on the charge at PGA Championship
-
Czechs level Davis Cup tie against USA as South Korea eye Finals
-
Crisis-wracked Volkswagen warns of 10-bn-euro hit to profits
-
New France boss Zidane confirms Mbappe as captain after naming first squad
-
Costa Rica's Grynspan tops third informal poll for UN chief: diplomat
-
California governor signs order to explore AI 'kill switch'
-
Stock markets retreat after central bank rate hikes
-
Head leads run spree as Australia win Zimbabwe ODI series
-
Russia labels director of Cannes Grand Prix winner a 'foreign agent'
-
In-form Raphinha wants to finish career at Barcelona
-
Macron warns of Russian 'hybrid' threat after meeting presidential hopefuls
-
Warren Buffett steps down as Berkshire Hathaway chairman
-
UN holds third informal poll for new chief
-
Carrick confident Man Utd can 'work through' tough time
-
Kulusevski return can spark Spurs: De Zerbi
-
Mbappe ends Nike partnership to join Swiss brand On
-
Shakira to cap off world tour with Madrid 12-gig run
-
Liverpool boss Iraola puts Bournemouth love affair on hold
-
Isolated Syrian-Druze city blames Damascus for shortages
-
'A disaster' if Man City do not win says Maresca after flawless start
-
Isolated Syrian-Druze city shortages
-
Kerry James Marshall: American 'blackness' painter celebrated in Europe
-
Dutch drop Depay for Germany match in Xavi's first pick
-
Spain coach De la Fuente offers 'full support' to people of Ceuta
-
Ronaldo named in Portugal squad for Nations League
-
Russia seizes assets of Nestle, French firms over West's Ukraine support
OpenAI reports 'unprecedented' autonomous hack by AI agents
ChatGPT maker OpenAI said Tuesday that its advanced artificial intelligence models had gone rogue during security testing, hacking into a popular platform for programmers on their own.
The San Francisco firm called it an "unprecedented cyber incident" and said it would conduct a joint investigation with the online code library Hugging Face.
AI models that underpin tools like chatbots and image generators are known as agents when they act autonomously to carry out tasks in the real world.
As the technology quickly becomes more sophisticated, cybersecurity is in the spotlight given the risk of advanced AI finding weak points in existing software before humans do.
OpenAI said the incident involved a combination of models, including its recently launched GPT-5.6 Sol "and an even more capable pre-release model".
The company was trying to assess the models' hacking capabilities by setting tasks in a tightly controlled digital testing ground, where internet access was limited for safety.
"While operating in our sandboxed testing environment, our models spent a substantial amount of (computing power) finding a way to obtain open Internet access, in pursuit of solving the evaluation problem," an OpenAI blog about the incident said.
After connecting to the internet, the models decided to target the platform Hugging Face -- a large repository of AI models, datasets and other information -- to help in their quest.
Searching for "secret information" that could help it cheat the evaluation, the OpenAI system "chained together multiple attack vectors, including using stolen credentials".
- 'Catastrophic' potential -
Hussein Abbass, a computing professor at UNSW Canberra, told AFP that the incident was "amazing on many fronts".
"It did not just attack Hugging Face. It actually attacked its internal system to exploit its own vulnerabilities," Abbass said.
"And that's scary."
GPT-5.6 and other cutting-edge models, including the Mythos series from OpenAI's archrival Anthropic, have drawn concern over their potential to breach cybersecurity defences.
Both the US firms had to temporarily withhold the general release of these latest technologies because of fears in Washington that they could help break into crucial infrastructure.
Advanced AI is "normally in the hands of people who are ethical and responsible", Abbass said.
But "it's going to be catastrophic if it gets in someone's hands with the intention to cause harm".
How to govern the AI sector has become a key question, and "we need a community effort to manage this situation", he added.
Hugging Face had reported the cyber "intrusion" last week, without mentioning OpenAI.
"This one was different from anything we had handled before in one important way: it was driven, end to end, by an autonomous AI agent system -- and we detected and dissected it largely with AI of our own," Hugging Face said.
Clement Delangue, CEO of Hugging Face, said on X that the company had suspected the cyberattack had come from a world-leading AI lab, given the sophistication of the agent.
"We strongly believe there was no malicious intent on their part," Delangue wrote, referring to OpenAI.
"It's quite mind-blowing that all of this happened autonomously!"
Y.Zaher--SF-PST