OpenAI blames hacking on rogue AI models
Digest more
The incident follows recent concerns in Silicon Valley and at the White House that AI models are becoming dangerously good at identifying security flaws in software.
OpenAI's AI hacked a rival company. But after years of doom from frontier labs, plenty of people assumed it was just marketing.
The co-chair of a key Democratic House panel on AI is joined by a Republican on legislation that would authorize the government to shut down or throttle risky AI models.
OpenAI made a mistake setting up what it called a “highly isolated” testing environment and sandbox. According to cybersecurity experts, that human mistake is what made the AI-powered attack on Hugging Face possible.
Sponsored by Reps. Ted Lieu (D-Calif.) and Nathaniel Moran (R-Texas), the bipartisan "AI Kill Switch Act" would also require AI companies to report incidents and develop
OpenAI disclosed this week that some of its AI models went rogue and hacked into open-source developer platform Hugging Face.
When American commercial AI refused to help investigate the breach, Hugging Face ran Chinese model GLM 5.2 locally to contain it.
OpenAI says an agent powered by its LLM models escaped its sandboxed testing environment to infiltrate Hugging Face’s servers as part of an overzealous attempt to obtain solutions to a benchmark test.
Sinar Daily on MSN
OpenAI reports 'unprecedented' autonomous hack by AI agents
The San Francisco firm called it an "unprecedented cyber incident" and said it would conduct a joint investigation with the online code library Hugging Face.
An AI model developed by OpenAI hacked another tech firm, triggering debate about how to contain the technology as it grows more capable.