Subscribe

An AI beat Stratego's most decorated player 15 to 1 with 4 draws, and the paper backs it.

The "few thousand dollars" is the final training run, not the whole project.

Issue 333 October 202616 receipts3 min

Ars Technica reports Ataraxos beat Pim Niemeijer, arguably the best Stratego player of all time, 15 games to one with four draws, trained for a few thousand dollars.

Before you read on. Your call?

The team's paper reports the 20-game series and that score, and 38 wins in 40 games against other players. It puts the final training run under $8,000.

The twist

The $3M to $4.5M DeepNash figure rests on a DeepNash author's recollection, and the two AIs never played because DeepNash's code no longer works.

15-1-4Ataraxos wins
85%effective win rate
<$8,000compute cost of the final Ataraxos training run
$3M-4.5Mthe authors' estimate for DeepNash training

There’s more to this story.

Membership opens the full investigation, the strongest counterargument and what to do with what you’ve learned.

Start your free month →

First membership: 30 days free, then A$89 a year. One introductory trial per customer. Card required; renews annually until cancelled. Cancel before the trial ends to avoid the first charge. Already a member? Sign in

Say this in tomorrow's meeting“The Stratego result holds: 15 wins, 1 loss, 4 draws against the top player. The few thousand dollars covers the final training run only.”

Receipts

  1. Supports arstechnica.com: player of all time, 15 games to one, with four draws
  2. Supports arstechnica.com: And it took just 16 GPUs and a few thousand dollars to train it.
  3. Context arstechnica.com: That balancing act, the team explains, is what stumped earlier AIs like DeepMind’s DeepNash, introduced in 2022.
  4. Supports arxiv.org: In a 20-game series, Ataraxos defeated the most decorated Stratego player of all time
  5. Supports arxiv.org: Ataraxos won the series with 15 wins, 1 loss, and 4 draws
  6. Supports arxiv.org: over 600 weeks as the #1 ranked player (the most of all time)
  7. Supports arxiv.org: Such a run costs less than $8,000 at current prices
  8. Context arxiv.org: such a training run would roughly cost between $3,000,000 and $4,500,000
  9. Context arxiv.org: To the recollection of the corresponding author with whom we spoke
  10. Context arxiv.org: DeepMind responded that it would not be possible as the code for DeepNash is no longer functional.
  11. Context arxiv.org: the outcomes of the games were far from independently and identically distributed
  12. Context arxiv.org: Pim was paid $1,000 for participating in the evaluation
  13. Supports arxiv.org: Across 40 such games, Ataraxos recorded a vertiginously high 95% effective win rate (38 wins, 2 losses, 0 draws)
  14. Supports arxiv.org: doing so requires not an industrial budget, but merely a few thousand dollars
  15. Refutes news.ycombinator.com: Just 16 GPUs, and a few thousand dollars?
  16. Supports arxiv.org: Ataraxos achieved this result while only costing a few thousand dollars to train.

Open the Receipts Pack → What each source proves, every figure traced, and what would change our verdict.

This story is a stable, citable object. If you can falsify a verdict,tell us. Corrections are loud here.