Sol Ultrafast finished Humanity's Last Exam in 11 hours. Whether it is still the same Sol remains unexamined.
The 750 tokens a second is probably real silicon. The quality-parity line has wiggle room its authors chose, the exam race was a batch job, and the price is a secret.
"GPT-5.6 Sol Ultrafast, served on Cerebras hardware, delivers up to 750 output tokens per second, 11x faster than Fable 5, finishing all of Humanity's Last Exam in 11 hours 11 minutes versus Fable 5's 78 hours 27 minutes, without any quality compromise." [SOURCE ↗]

THE CLAIM. OpenAI and Cerebras say GPT-5.6 Sol in Ultrafast mode hits up to 750 output tokens per second, 11x faster than Fable 5, and blitzed all 2,500 Humanity's Last Exam questions in 11 hours 11 minutes against Fable 5's 78 hours 27 minutes, with no quality compromise.
THE CHECK. the silicon is plausible and the framing is theater. Answering 2,500 independent questions is an embarrassingly parallel workload, so the wall-clock race measures cluster scale, not model speed, and per-question latency is not disclosed. Neither company states that Ultrafast performs identically to regular Sol, a sentence they would shout if they could write it, and the industry's record on 'no quality loss' serving modes is poor. No price, no context limit, no configuration. The fine print concedes results may vary by workload and configuration.
Cerebras and OpenAI jointly announced GPT-5.6 Sol Ultrafast, a serving mode for OpenAI's flagship on Cerebras wafer-scale hardware, claiming up to 750 output tokens per second, 11x faster than Fable 5 and 5x faster than Opus 4.8 on Fast mode, with comparisons drawn from speeds reported by Artificial
🔒 THE FULL AUTOPSY · FREE WITH AN ACCOUNTYou just read the free check. Sign in free, a code by email, no passwords, and the rest unlocks: the evidence trail, the steelman and the rebuttal, all 7 sources with quotes and screenshots, and our on-record call.
Couldn't verify your access — this looks like our error, not yours.