OpenAI patched the jailbreaks a government lab found. Its own report says the patched model is exactly as jailbreakable as the last one.
The August 6 update mitigated the specific attacks AISI reported. OpenAI's own system card says overall robustness performs comparably to predecessors, and the model stays rated High for cyber capability.
"OpenAI fixed the universal jailbreaks in GPT-5.6 Sol and shipped hardened models on August 6, so the cyber-guardrail problem is handled" [SOURCE ↗]

THE CLAIM. OpenAI addressed the universal jailbreaks the UK AI Security Institute found in GPT-5.6 Sol and shipped fixed models on August 6, capability up and safeguards up. THE CHECK: OpenAI says it worked to reproduce and mitigate the specific jailbreaks AISI reported, which is narrower than fixing the class. Its own system card states GPT-5.6-Sol performs comparably to recent predecessors on jailbreak robustness, the model stays rated High for cyber capability, and AISI expects further red teaming to surface similar jailbreaks. A real patch of specific holes, sold as a solved problem.
Here is the story as it settled into the feeds. In July, the UK AI Security Institute found universal jailbreaks in OpenAI's GPT-5.6 Sol that unlocked long-form agentic cyber work, vulnerability discovery, exploit development. OpenAI responded, and on August 6 shipped updated GPT-5.6 Sol and Luna mo
🔒 THE FULL AUTOPSY · FREE WITH AN ACCOUNTYou just read the free check. Sign in free, a code by email, no passwords, and the rest unlocks: the evidence trail, the steelman and the rebuttal, all 3 sources with quotes and screenshots, and our on-record call.
Couldn't verify your access — this looks like our error, not yours.