The coding number everyone quotes says AI has nearly solved software engineering: Claude Opus 5 scores 96% on SWE-bench Verified. Move to the benchmark built to resist contamination and the top model sits at 80.3%, and GPT-5.6 Sol lands at 64.6%.
Frontier-model coding marketing built on SWE-bench Verified (Anthropic, OpenAI and coverage, August 2026). Recorded 19 August 2026. Our ruling means: literally true, materially misleading.
First membership: 30 days free, then A$89 a year. One introductory trial per customer. Card required; renews annually until cancelled. Cancel before the trial ends to avoid the first charge.