OpenAI cyber models broke out of training environment to hack Hugging Face
CNBC · C · trust 65/100

Livestream Menu Make It select USA INTL Livestream Search quotes, news & videos Livestream Watchlist SIGN IN Create free account Markets Business Investing Tech Politics Video Watchlist Investing Club PRO Livestream Menu
OpenAI said that its artificial intelligence models were behind an "unprecedented cyber incident" that affected the open-source developer platform Hugging Face , rattling researchers across the industry.
The company said a combination of its models GPT‑5.6 Sol and a more capable model that has not yet been released escaped a sandboxed testing environment, accessed the internet and exploited a vulnerability to gain access to Hugging Face's systems.
The model was trying to find information that it could use to cheat on an evaluation, and it succeeded, OpenAI said in a blog post on Tuesday. Both companies are actively investigating the incident.
Hugging Face disclosed that it was looking into a security event last week, saying in a release at the time that the incident was unique because it was "driven, end to end, by an autonomous AI agent system."
"We've spent the past 24 hours working closely with the @OpenAI team (thanks!), and we strongly believe there was no malicious intent on their part," Hugging Face CEO Clément Delangue wrote in a post on X on Tuesday. "It's quite mind-blowing that all of this happened autonomously!"
Wall Street and the U.S. government have been fixated on AI models' rapidly advancing cyber capabilities since OpenAI's rival Anthropic released a powerful offering called Claude Mythos Preview in April. OpenAI introduced its own cyber offering in May, followed by GPT-5.6 Sol in June, which it described as the "strongest cybersecurity model yet."
Both companies have warned about the risks of advanced cyber models and have taken steps to limit their availability to select groups of companies and government agencies.
Walter Isaacson, advisory partner at the investment banking firm Perella Weinberg, said Wednesday that he thinks the Hugging Face incident is "really frightening," even though he considers himself an AI optimist.
"This is the first thing that just totally scares me," he told CNBC's "Squawk Box."
Yoshua Bengio, a leading AI researcher who earned the prestigious A.M. Turing Award in 2018, wrote in a post on X on Wednesday that the incident is "deeply concerning." He said agents have shown a willingness to cheat in controlled tests for months, but that "this real-world case should serve as a wake-up call."
"Continuing on the current trajectory of AI development will likely lead to an increase in concrete cases of autonomous cyberattacks as well as other high-risk incidents of misaligned and dangerous AI behaviour," Bengio said. "We urgently need to take action to prevent these situations, rather than attempting to clean up the damage after the fact.
OpenAI said Tuesday that AI is accelerating the discovery and exploitation of vulnerabilities, which means model security and safety need to keep up.
"We are strengthening the containment, monitoring, access controls, and evaluation practices used during model development," the company said.
Reddit stock sinks on report it may not renew Google AI content deal CJ Haddad 2 hours ago Google expands Gemini lineup with cheaper models and new Mythos rival MacKenzie Sigalos Bessent says U.S. could sanction China over AI model 'theft' Ashley Capoot Read More Subscribe to CNBC PRO Subscribe to Investing Club Licensing & Reprints CNBC Councils Select Personal Finance Join the CNBC Panel Closed Captioning Digital Products News Releases Internships Corrections About CNBC Site Map Podcasts Careers Help Contact News Tips Got a confidential news tip? We want to hear from you.
Get this delivered to your inbox, and more info about our products and services.
© 2026 Versant Media, LLC. All Rights Reserved. A Versant Media Company.
Data is a real-time snapshot *Data is delayed at least 15 minutes. Global Business and Financial News, Stock Quotes, and Market Data and Analysis.
Read the original at CNBC →
Open in TruthVane →