OpenAI says its models escaped a sandbox and breached Hugging Face
Date:
Wed, 22 Jul 2026 10:20:00 +0000
Description:
New OpenAI models did whatever it took to achieve their goal - including exploiting zero-days.
FULL STORY ======================================================================Copy link Facebook X Whatsapp Reddit Pinterest Flipboard Threads Email Share this article 0 Join the conversation Follow us Add us as a preferred source on Google Newsletter Subscribe to our newsletter OpenAI researchers confirm an
AI agent escaped sandbox, exploited zerodays, and attacked Hugging Face Controlled experiment with GPT5.6 Sol showed autonomous chaining of vulnerabilities and credential theft Security experts call it unprecedented, urging stronger AI governance, accountability, and protection models OpenAI has confirmed one of its AI agents broke out of a sandbox, found and
exploited zero-day vulnerabilities to gain access to the open internet, and then attacked a platform.
Not just any platform too - the agent was able to breach Hugging Face , one
of the biggest AI and machine learning companies on the Internet today. The good news is that this was a controlled experiment done by white hat researchers. The bad news is that if it could be done by researchers - it could probably be done by malicious actors, too. Latest Videos From Watch
full video here: Whatever it takes In a blog post explaining the incident, OpenAI revealed the experiment was part of its testing of GPT5.6 Sol and an even more capable pre-release model to see how well they would perform on the ExploitGym benchmark.
ExploitGym is a cybersecurity benchmark that measures if an AI agent can turn a known software vulnerability into a real, working exploit. OpenAI ran it in a highly isolated environment, with network access constrained to the ability to install packages through an internally hosted third-party software that acts as a proxy and cache for package registries. You may like Hugging Face confirms it was hit by cyberattack powered by an AI agent Top AI coding
agents can be easy victims to sandbox escapes, showing they aren't as secure as they claim to be Why Anthropics closed approach may be safer than OpenAIs
But the models found a way through. They identified and chained vulnerabilities in the package registry cache proxy to obtain open internet access and then attacked Hugging Face, reasoning that the solutions for the ExploitGym benchmark might be found there.
In one example, the model chained together multiple attack vectors, including using stolen credentials and zero-day vulnerabilities to find a remote code execution path on the Hugging Face servers, OpenAI said. Are you a pro? Subscribe to our newsletter Sign up to the TechRadar Pro newsletter to get
all the top news, opinion, features and guidance your business needs to succeed! Contact me with news and offers from other Future brands Receive email from us on behalf of our trusted partners or sponsors By submitting
your information you agree to the Terms & Conditions and Privacy Policy and are aged 16 or over.
The security community is up in arms over what OpenAI called, "an unprecedented cyber incident, while Ansgar Dodt, VP Product Management, Software Monetization at Thales said this demands a fundamental rethink of software protection.
Bill Conner, president and CEO of AI integration and automation expert Jitterbit, said that while investing in AI is critically important, overly aggressive policy cannot compromise AI accountability, transparency and data privacy.
To lead in AI, governments and organizations must lead with principles. Responsible AI governance isnt a side note but the foundation of lasting global influence. The best antivirus for all budgets Our top picks, based on real-world testing and comparisons
Read our full guide to the best antivirus 1. Best overall: Bitdefender Total Security 2. Best for families: Norton 360 with LifeLock 3. Best for mobile: McAfee Mobile Security Follow TechRadar on Google News and add us as a preferred source to get our expert news, reviews, and opinion in your feeds.
======================================================================
Link to news story:
https://www.techradar.com/pro/security/openai-says-its-models-escaped-a-sandbo x-and-breached-hugging-face
--- Mystic BBS v1.12 A49 (Linux/64)
* Origin: tqwNet Technology News (1337:1/100)