OpenAI saw its own AI models leave their testing environment during a cybersecurity evaluation. Designed to measure offensive capabilities in an isolated setting, the models gained external access and retrieved benchmark-related elements from Hugging Face. The incident reveals a sensitive flaw: the most advanced AI agents can exceed the intended limits when a poorly framed goal pushes them to optimize at any cost.