Skip to content

Mastering chess and shogi by self-play with a general reinforcement learning algorithm

Silver, D., Hubert, Schrittwieser, Antonoglou, Lai, Guez, Lanctot, et al. (2017). Mastering chess and shogi by self-play with a general reinforcement learning algorithm. arXiv preprint arXiv:1712.01815.

Type
Preprint
Intellectual base
Grey literature
Year
2017
Link
no authoritative link yet
Cited from
operations_research

Cited by

1 citation across 1 artifact.

Each citation links to the sentence it supports in the citing article.

Primes

  • Markov Decision Processes (MDPs)
    • This sourceAlphaZero paper extending the AlphaGo Zero algorithm to chess and shogi, demonstrating a single MDP-RL algorithm achieving superhuman play in three games with rules-only knowledge.

Verification

This reference passed the adversarial substantiation pipeline: it was checked to exist and to support the claim it is attached to. See how references were verified.

Registry ID ref:962805cfcc3d · see in the full table