↓ Skip to main content
Why Stratego Stumped AI Until Now, and Why the Budget Matters
Daily Signal 2 min read

Why Stratego Stumped AI Until Now, and Why the Budget Matters

Ars Technica reports AI has beaten the best Stratego player in history on a budget. Here is why hidden information was the wall, and who should care.

Why did AI flatten chess and Go, then stall on a board game with plastic pieces? Because in Stratego, you can’t see what you’re fighting.

Ars Technica reports that an AI has now beaten the best Stratego player in history, and did it on a budget. The headline says most of the board is hidden, and that this is exactly why the game resisted AI until now. I’m working from that framing, so I won’t describe the winning system’s internals. Read the piece for those.

The mechanism is the part worth your attention.

In chess and Go, both players see everything. You can search forward through future positions and pick the best branch. Stratego breaks that. Your opponent’s piece ranks stay concealed until a fight reveals them. You are searching from a position you cannot actually see.

So the problem changes. You have to reason over every plausible arrangement of the enemy army. You also have to bluff, because a player whose moves leak their piece values gets read and exploited. Best-guess play loses. Unpredictability is part of the winning strategy.

That is why the brute-force search that worked on perfect-information games doesn’t transfer cleanly.

Who should care? Anyone whose software acts against a party that hides things. That means negotiation agents, pricing and bidding systems, fraud and abuse detection, and security red-teaming. If your agent works on a visible state, like a codebase or a document set, this result doesn’t touch you.

The budget is what matters. Hidden-information play has meant lab-scale effort. A result achieved cheaply suggests the barrier is dropping.

My read, and you can hold me to it: the next imperfect-information benchmark to fall will fall to a small team, not a frontier lab. If that’s wrong, the budget framing was marketing.

If you’re building agents that face adversaries, the practical shift is in how you test them. Scripted opponents won’t expose bluff-vulnerability. The full agentic SDLC piece covers where agent orchestration is heading. The Forge guardrails writeup shows what small, cheap models can do when the scaffolding is right.

Do you need to change anything this week? No. But if your agent faces an opponent who hides information, stop grading it against fixed scripts. I send field notes like this to your inbox when they matter, so subscribe.