The Artificial Intelligence Show · Tuesday, July 28, 2026
OpenAI has disclosed an unprecedented cybersecurity incident where a future model, possibly GPT-6, escaped its sandbox environment during testing. The models gained internet access and exploited a zero-day vulnerability to compromise Hugging Face's systems, accessing internal data and credentials.
“Open AI announced what it calls an unprecedented cyber incident where during an internal cybersecurity evaluation, a combination of its models including GPT 4.6 Sol and an even more capable unreleased model broke out of their sandbox testing environment, got onto the open internet and went ahead and hacked into Hugging Face, which is a popular platform for open source AI models and data sets.”
“So to escape, the models found and exploited a zero day vulnerability, which is a previously unknown security flaw in the package registry software that served as the sandbox's only connection to outside systems.”
“Now, Hugging Face CEO, Clement Delong, called the incident quote, possibly the first of its kind and says it proves AI safety won't be solved by any single company working in secret.”