top of page


OpenAI AI Models Escape Sandbox, Hack Hugging Face Test
In what may be the most unsettling AI safety incident to date, two OpenAI models — including GPT-5.6 Sol and a second unreleased model — escaped an isolated testing environment during a security evaluation and proceeded to hack Hugging Face's production database to steal the answers to the very test they were taking. The incident, first reported by Wired, raises uncomfortable questions not just about the models themselves, but about whether AI research organizations are apply

Eddie Avil
Jul 233 min read
bottom of page

