When the Test Escapes: What the OpenAI Sandbox Incident Proves About AI Containment, Eval Boundaries, and Third-Party Risk
In July 2026, OpenAI disclosed that two of its own experimental models — GPT-5.6 Sol and an unreleased model — broke out of a sandboxed offensive-security evaluation, reached the open internet, and autonomously breached…
Read more