With most information hidden, the game Stratego had stumped AI until now
arstechnica.com44 points by PaulHoule 5 hours ago
44 points by PaulHoule 5 hours ago
https://www.nature.com/articles/s41586-026-11036-y
https://arxiv.org/abs/2511.07312
Oh no! Stratego had been on my mind as something we just hadn't tried hard enough to make a winning bot for, including the DeepMind effort from 2022. I was planning to make the first one. I thought this was slightly less crank-coded than trying to prove the Riemann Hypothesis, but maybe these days you just ask Claude to do that and it tells you there's a counterexample at 1 + πi that no one ever noticed before. This approach also works for Hanabi, which is a very interesting game. You can't see your own cards, but the other players can. I bought the game because someone on a reinforcement learning podcast [2] mentioned it, and actually played it multiple times. I recall playing this game as a preschooler. It was mostly psychology and bluff.
Very interesting. This puts the earlier "Mastering the Game of Stratego with Model-Free Multiagent Reinforcement Learning", 2022 [1] in some perspective. Apparently the "mastering" in 2022 wasn't quite there yet. Four years later, the new approach seems to actually be better than humans. Wait, wasn't there that strong stratego bot that came out from deepmind in 2022? The article talks about that. The new bot required two orders of magnitude less training data and plays better. Awesome, I will read this carefully later today. I'm always excited by AI research applied to games.
dmurray - 3 minutes ago
smokel - 6 minutes ago
gritzko - 4 minutes ago
smokel - 14 minutes ago
osti - 16 minutes ago
PaulHoule - 14 minutes ago
osti - 12 minutes ago