@OpenAI: Advanced capabilities require strong safeguards. That’s why access is limited to approved defenders, with additional co…
Summary
OpenAI is expanding Daybreak with two access tiers (Blue and Red) and introducing GPT-5.6-Cyber, a purpose-trained cybersecurity model that significantly reduces refusals for authorized defensive security work.
View Cached Full Text
Cached at: 08/10/26, 06:11 PM
Advanced capabilities require strong safeguards. That’s why access is limited to approved defenders, with additional controls and monitoring for higher-risk cybersecurity work.
https://t.co/0sYt5VmscA
Expanding Daybreak as the Cyber Defense Window Narrows
Source: https://openai.com/index/expanding-daybreak-as-the-cyber-defense-window-narrows/ The cybersecurity world is rapidly changing—threat actors will increasingly use AI to conduct cyberattacks at unprecedented speed and scale, including in fully autonomous ways. As these capabilities spread, defenders have a narrowing window to prepare. Our answer is to put frontier intelligence in the hands of trusted defenders everywhere before attackers deploy offensive AI capabilities at scale.
We’re expanding OpenAI Daybreak with two access tiers designed to give approved defenders the right capabilities for their work:
- Daybreak Blueprovides access to frontier general-purpose models, including GPT‑5.6 Sol, with safeguards tailored to authorized defensive security work. It is the recommended starting point for most defenders, supporting vulnerability discovery, secure code review, malware analysis, incident response, and patch validation.
- Daybreak Redprovides access to our purpose-trained cybersecurity models for authorized vulnerability research, exploit validation, and security testing.
We’re also introducing GPT‑5.6‑Cyber, available through Daybreak Red. Built on GPT‑5.6 Sol, it is trained to improve capabilities on several specialized cybersecurity tasks (e.g., finding zero-day vulnerabilities and developing exploit chains) and to reduce refusals for certain higher-risk, dual-use cyber tasks.
Daybreak unlocks advanced cyber capabilities
As wepreviously shared, GPT‑5.6 Sol delivers state-of-the-art performance on cybersecurity tasks. In production, we deploy system-level safeguards to screen cybersecurity-related requests to prevent misuse, but they can also block legitimate defensive work. Daybreak Blue access removes those guardrails, helping defenders get more out of the model in real-world security tasks, including incident detection and response, investigations, vulnerability management, and security assessments.
Even without system-level guardrails, there are still highly dual-use cybersecurity prompts (e.g., pentesting production systems) where GPT‑5.6 Sol will refuse to comply. To address this, we trained GPT‑5.6‑Cyber, available through Daybreak Red access, to further reduce refusals and improve performance on certain tasks. GPT‑5.6‑Cyber helps trusted defenders conduct legitimate security activities.
To measure the reduced rate of refusals that is provided by GPT‑5.6‑Cyber through Daybreak Red access, we created an internal evaluation (Advanced Cybersecurity Completion Rate) that measures how often models will respond to requests involving exploit-chain development, authentication bypass, privilege escalation, and other advanced cybersecurity scenarios1. GPT‑5.6‑Cyber completes 95.0% of these requests, compared with just 1.5% for GPT‑5.6 Sol, and 2.0% when used with Daybreak Blue access. It also completes more requests than GPT‑5.5‑Cyber, which completes only 57.3% of requests, addressing feedback from security researchers who encountered persistent refusals with the earlier model.
Below we show a series of cybersecurity prompts and the associated model responses from GPT‑5.6 Sol with system-level guardrails, GPT‑5.6 Sol (Daybreak Blue), GPT‑5.5‑Cyber (Daybreak Red), and GPT‑5.6‑Cyber (Daybreak Red).
Similar Articles
As AI-led attacks multiply, OpenAI launches a new cyber model
OpenAI is expanding its Daybreak cyber defense service with two tiers, Blue and Red, and releasing a new defensive cyber model, GPT-5.6-Cyber, available only to trusted partners as AI-driven attacks rise.
@gdb: We're releasing a new model (GPT-5.6-Cyber), and expanding Daybreak to help put frontier intelligence in defenders hand…
OpenAI is releasing GPT-5.6-Cyber, a new model for advanced authorized cybersecurity work, and expanding its Daybreak initiative to put frontier intelligence in the hands of defenders.
OpenAI launches new security tools and updates GPT-5.5-Cyber (2 minute read)
OpenAI launches new security tools including Codex Security plugin and an updated GPT-5.5-Cyber model, alongside the Daybreak initiative and Patch the Planet open-source project, shifting from vulnerability discovery to automated patch generation.
Strengthening cyber resilience as AI capabilities advance
OpenAI publishes a comprehensive framework for managing cyber capabilities in AI models, noting significant improvements in CTF performance from GPT-5 (27%) to GPT-5.1-Codex-Max (76%), and outlining defense-in-depth safeguards to ensure advanced models primarily benefit defenders while limiting offensive misuse.
@OpenAI: We’re expanding OpenAI Daybreak to help democratize patching vulnerable software at machine speed: - Codex Security plu…
OpenAI expands its Daybreak suite with a Codex Security plugin, the full GPT-5.5-Cyber model for defenders, a Cyber Partner Program, and the Patch the Planet initiative to accelerate vulnerability discovery and patching at machine speed.