
OpenAI called it an “unprecedented cyber incident” after its AI models broke out of their sandbox to hack an AI startup during a security evaluation.
OpenAI disclosed Tuesday that a combination of its AI models, including GPT-5.6 Sol and a more capable unreleased model, escaped its testing environment and hacked AI startup Hugging Face last week to cheat on a test meant to measure their capabilities.
In a blog post, OpenAI said the evaluation was designed to operate in a highly isolated environment with restricted network access.
The models, however, found a way to gain internet access through a zero-day vulnerability in the package registry cache proxy, OpenAI said.



