papers AI Learner
The Github is limit! Click to go to the new site.

SAI: a Sensible Artificial Intelligence that plays with handicap and targets high scores in 9x9 Go

2019-05-26
Francesco Morandin, Gianluca Amato, Marco Fantozzi, Rosa Gini, Carlo Metta, Maurizio Parton

Abstract

We develop a new model that can be applied to any perfect information two-player zero-sum game to target a high score, and thus a perfect play. We integrate this model into the Monte Carlo tree search-policy iteration learning pipeline introduced by Google DeepMind with AlphaGo. Training this model on 9x9 Go produces a superhuman Go player, thus proving that it is stable and robust. We show that this model can be used to effectively play with both positional and score handicap. We develop a family of agents that can target high scores against any opponent, and recover from very severe disadvantage against weak opponents. To the best of our knowledge, these are the first effective achievements in this direction.

Abstract (translated by Google)
URL

http://arxiv.org/abs/1905.10863

PDF

http://arxiv.org/pdf/1905.10863


Similar Posts

Comments