With most information hidden, the game Stratego
摘要
卡内基梅隆、MIT、纽约大学与斯坦福的研究团队开发出AI“Ataraxos”,在Stratego(战略棋)中以15胜1负4平击败顶尖人类选手Pim Niemeijer。Stratego属于不完全信息博弈,棋子身份仅在碰撞时揭示,隐藏信息量大且随时间逐步展开。此前DeepMind等机构未能造出稳定击败顶尖人类玩家的机器。该AI训练仅用16块GPU,成本数千美元
Deep Blue took down Garry Kasparov at chess in 1997, AlphaGo beat Lee Sedol at Go in 2016, and poker bots have been beating professionals for years. But one classic game called Stratego held out. Even DeepMind, with its exceptional budget, couldn't build a machine that reliably beat the best human players.
Now, a team of researchers from Carnegie Mellon, MIT, New York University, and Stanford University has done it. Their AI, called Ataraxos, beat Pim Niemeijer, arguably the best Stratego player of all time, 15 games to one, with four draws. And it took just 16 GPUs and a few thousand dollars to train it.
Hidden armies
In Stratego, each player gets 40 pieces representing military ranks, from a marshal down to a spy, plus bombs and a flag. You win by capturing the opponent's flag. Your opponent knows where your pieces are, but not what they are. Identities are revealed only when two pieces collide in battle—the weaker one is removed, and the identity of the winner is revealed. That makes Stratego an imperfect-information game, just like poker, which computers cracked years ago. “There's something super distinctive about Stratego, which is that it is a massive amount of hidden information that unfolds over a very long time scale,” said Eugene Vinitsky, a researcher at NYU and co-author of the study.
转载信息
评论 (0)
暂无评论,来留下第一条评论吧