GET THE AUTOPSY ➔

Sol Ultrafast finished Humanity's Last Exam in 11 hours. Whether it is still the same Sol remains unexamined.

The 750 tokens a second is probably real silicon. The quality-parity line has wiggle room its authors chose, the exam race was a batch job, and the price is a secret.

01THE CLAIM
"GPT-5.6 Sol Ultrafast, served on Cerebras hardware, delivers up to 750 output tokens per second, 11x faster than Fable 5, finishing all of Humanity's Last Exam in 11 hours 11 minutes versus Fable 5's 78 hours 27 minutes, without any quality compromise." [SOURCE ↗]
TRUE, BUT7 SOURCES · LIVE 2026-08-25
CEREBRAS + OPENAI TRACK RECORD1 CLAIM · 40/100 BS RATE →
750TOK/S, 'UP TO', CONFIG UNDISCLOSED
62.2WHAT API SOL MEASURES ON THE OPEN API, PER AA
0PRICES, CONTEXT LIMITS, OR CONFIGS PUBLISHED
Sol Ultrafast finished Humanity's Last Exam in 11 hours. Whether it is still the same Sol remains unexamined.
02THE CHECK

THE CLAIM. OpenAI and Cerebras say GPT-5.6 Sol in Ultrafast mode hits up to 750 output tokens per second, 11x faster than Fable 5, and blitzed all 2,500 Humanity's Last Exam questions in 11 hours 11 minutes against Fable 5's 78 hours 27 minutes, with no quality compromise.

THE CHECK. the silicon is plausible and the framing is theater. Answering 2,500 independent questions is an embarrassingly parallel workload, so the wall-clock race measures cluster scale, not model speed, and per-question latency is not disclosed. Neither company states that Ultrafast performs identically to regular Sol, a sentence they would shout if they could write it, and the industry's record on 'no quality loss' serving modes is poor. No price, no context limit, no configuration. The fine print concedes results may vary by workload and configuration.

03SAY THIS IN THE MEETING · 📸 SCREENSHOT IT
"The fast is probably real, the parity is asserted. Until there is per-question latency, a price, and a full eval suite on Ultrafast, treat it as a different product."

Cerebras and OpenAI jointly announced GPT-5.6 Sol Ultrafast, a serving mode for OpenAI's flagship on Cerebras wafer-scale hardware, claiming up to 750 output tokens per second, 11x faster than Fable 5 and 5x faster than Opus 4.8 on Fast mode, with comparisons drawn from speeds reported by Artificial

🔒 THE FULL AUTOPSY · FREE WITH AN ACCOUNT

You just read the free check. Sign in free, a code by email, no passwords, and the rest unlocks: the evidence trail, the steelman and the rebuttal, all 7 sources with quotes and screenshots, and our on-record call.

This story is a stable, citable object. If you can falsify a verdict, tell us. Corrections are loud here.