Mastering chess and shogi by self-play with a general reinforcement learning algorithm¶
Silver, D., Hubert, Schrittwieser, Antonoglou, Lai, Guez, Lanctot, et al. (2017). Mastering chess and shogi by self-play with a general reinforcement learning algorithm. arXiv preprint arXiv:1712.01815.
Cited by¶
1 citation across 1 artifact.
Each citation links to the sentence it supports in the citing article.
Primes¶
- Markov Decision Processes (MDPs)
This sourceAlphaZero paper extending the AlphaGo Zero algorithm to chess and shogi, demonstrating a single MDP-RL algorithm achieving superhuman play in three games with rules-only knowledge.
Verification¶
This reference passed the adversarial substantiation pipeline: it was checked to exist and to support the claim it is attached to. See how references were verified.
Registry ID ref:962805cfcc3d · see in the full table