Hacker News

PaulHoule
With most information hidden, the game Stratego had stumped AI until now arstechnica.com

https://www.nature.com/articles/s41586-026-11036-y

https://arxiv.org/abs/2511.07312


dmurray3 minutes ago

Oh no! Stratego had been on my mind as something we just hadn't tried hard enough to make a winning bot for, including the DeepMind effort from 2022. I was planning to make the first one.

I thought this was slightly less crank-coded than trying to prove the Riemann Hypothesis, but maybe these days you just ask Claude to do that and it tells you there's a counterexample at 1 + πi that no one ever noticed before.

smokel6 minutes ago

This approach also works for Hanabi, which is a very interesting game. You can't see your own cards, but the other players can. I bought the game because someone on a reinforcement learning podcast [2] mentioned it, and actually played it multiple times.

[1] https://en.wikipedia.org/wiki/Hanabi_(card_game)

[2] https://www.talkrl.com/episodes/jakob-foerster

gritzko4 minutes ago

I recall playing this game as a preschooler. It was mostly psychology and bluff. Very interesting.

smokel14 minutes ago

This puts the earlier "Mastering the Game of Stratego with Model-Free Multiagent Reinforcement Learning", 2022 [1] in some perspective. Apparently the "mastering" in 2022 wasn't quite there yet. Four years later, the new approach seems to actually be better than humans.

[1] https://arxiv.org/abs/2206.15378

osti16 minutes ago

Wait, wasn't there that strong stratego bot that came out from deepmind in 2022?

PaulHouleop14 minutes ago

The article talks about that. The new bot required two orders of magnitude less training data and plays better.

osti12 minutes ago

Awesome, I will read this carefully later today. I'm always excited by AI research applied to games.

bananaflag22 minutes ago

So cute, like some news story from 2019.

hn-front (c) 2024 voximity
source