SUBSCRIBE

OpenAI hit the brakes on one cyber model and sold another one three days later.

Astra might be too dangerous to release. GPT-5.6-Cyber, trained to refuse less, is for sale now.

01THE CLAIM
"We cannot rule out critical cyber capabilities under our Preparedness Framework" [SOURCE ↗]

THE MOVE: ZERO UNDERNEATH, the headline number has nothing behind it

TRUE, BUT6 SOURCES · LIVE 2026-09-05
OPENAI TRACK RECORD37 CLAIMS · 39/100 BS RATE →
0EVAL SCORES PUBLISHED FOR ASTRA
95.0%GPT-5.6-CYBER EXPLOIT COMPLETION
1.5%SAME EVAL, BASE GPT-5.6 SOL
3DAYS BETWEEN BRAKES AND GAS
OpenAI hit the brakes on one cyber model and sold another one three days later.
02THE CHECK

THE CLAIM. OpenAI said on August 7 it cannot rule out that its unreleased Astra model has 'critical' cyber capabilities, so it is slowing parts of the work.

THE CHECK. the post publishes the threshold definition and exactly zero measurements. No scores, no success rates, no trial counts. An analyst quoted by CSO Online calls it a precautionary trigger, not a finished finding. Media turned that into 'OpenAI pauses Astra'. OpenAI paused only internal activities that did not yet meet new security controls.

THE TWIST. three days later the same company expanded Daybreak and launched GPT-5.6-Cyber, trained to refuse less on 'higher-risk, dual-use cyber tasks'. It completes 95.0% of exploit-development requests, versus 1.5% for the base model. When OpenAI wants to publish numbers, it publishes numbers.

03SAY THIS IN THE MEETING
"When a lab says a model is too dangerous to ship, ask for the score. Astra got a press release with none. GPT-5.6-Cyber got a 95% next to a 1.5%."

On August 7, 2026, OpenAI published a post saying preliminary evaluations of its unreleased Astra model, concluded the night before, meant it could not rule out that Astra crosses the 'Critical' cybersecurity threshold in its Preparedness Framework. Under that framework, Critical means a model that

🔒 THE FULL AUTOPSY · FREE WITH AN ACCOUNT

You just read the free check. Sign in free, a code by email, no passwords, and the rest unlocks: the evidence trail, the steelman and the rebuttal, all 6 sources with quotes and screenshots, and our on-record call.

This story is a stable, citable object. If you can falsify a verdict, tell us. Corrections are loud here.