OpenAI has released a model built to do the offensive-security work its general-purpose models are trained to refuse.
GPT-5.6-Cyber, announced Monday, is trained to find zero-day vulnerabilities and develop exploit chains, and to "reduce refusals for certain higher-risk, dual-use cyber tasks."1 It is available only through Daybreak Red, an approval tier for vetted defenders.
The gap that training opens is wide. On an internal benchmark built around privilege escalation, bypassing authentication and building exploit chains, OpenAI says GPT-5.6-Cyber completes 95.0% of requests. GPT-5.6 Sol, the general-purpose model it is built on, completes 1.5%, and 2.0% with production guardrails removed. The previous cyber model, GPT-5.5-Cyber, completes 57.3%.1
Eric Wallace, an OpenAI researcher who co-leads its alignment training team, called it the company's "first large-scale attempt at directly improving capabilities for advanced cybersecurity tasks such as exploit development," and said OpenAI uses it internally for red-teaming.2
