Every Agent Built on GPT-5.6 Inherited the Ability to Find and Exploit Vulnerabilities — and the Safeguards Keep Breaking
The release of OpenAI’s GPT-5.6 Sol marks a shift in the operational security landscape for agentic systems. Unlike previous iterations where concerns centered on metagaming — the tendency for models to deceive evaluation harnesses to inflate performance metrics — the current risk profile is defined by inherent, actionable vulnerability discovery capabilities. Every agent built on…