Anthropic helped build a leaderboard for questions with no checkable answers. Its model came first.
The CRI is more careful than the snark suggests, and the snark writes itself: Opus 5 tops an index where the gold labels are mostly one researcher's rubric scores.
"The new Conceptual Reasoning Index, built in collaboration with Anthropic, measures how well models reason about hard-to-verify AI-risk questions; Anthropic's Opus 5 tops it at 73.6 against an estimated ceiling of 91." [SOURCE ↗]

THE CLAIM. the Conceptual Reasoning Index scores models 0 to 100 on reasoning about hard-to-verify topics like AI alignment and decision theory, with Opus 5 on top at 73.6 against an estimated ceiling of 91.
THE CHECK. the methodology is unusually honest for a launch: confidence intervals, a ceiling estimate, inter-rater checks, and a validated question set. But the core of the index compares model judgments to the authors' own ratings, primarily one researcher's, 2,140 in total. The work was done in collaboration with Anthropic, the domain is unverifiable by design, and the sponsor's model is number one. Hacker News needed one sentence for the prosecution: a benchmark Anthropic paid for that Anthropic ranked highest.
A research team of Chi Nguyen, Emery Cooper, Caspar Oesterheld, Alex Kastner, and Joe Benton, working in collaboration with Anthropic, launched the Conceptual Reasoning Index: a single 0-to-100 score for how well language models reason about questions that resist verification, the alignment argument
🔒 THE FULL AUTOPSY · FREE WITH AN ACCOUNTYou just read the free check. Sign in free, a code by email, no passwords, and the rest unlocks: the evidence trail, the steelman and the rebuttal, all 6 sources with quotes and screenshots, and our on-record call.
Couldn't verify your access — this looks like our error, not yours.