CANONICAL HISTORY
Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
David Silver and colleagues submitted Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm to arXiv on December 5, 2017. The paper described AlphaZero as a single algorithm that learned through self-play from random play with no domain knowledge beyond game rules and reported superhuman performance in chess, shogi, and Go.
Evidence / resource
This page preserves the public LINEAiGE record and its first-party source relationship.
Record identity
LINEAiGE IDalphazero-2017