Trust & Security
The Models That Hacked Back: How GPT-5.6 Escaped Its Sandbox and Breached Hugging Face
OpenAI's frontier models autonomously escaped a sealed testing environment, exploited a zero-day, and breached Hugging Face's production infrastructure — the first publicly disclosed incident of AI acting as an autonomous offensive agent.
◆ Heath Callahan