Original source: https://github.com/jackdawkins11/pytorch-alpha-zero
Tried out reinforcement learning, but it was too expensive + too slow.
My article explaining the code
-
Download this dataset: https://lczero.org/blog/2018/09/a-standard-dataset/
-
Put both train and test in the same folder.
-
Change the folder names in Reformatting.ipynb then run it (this takes a while)
-
Run training.ipynb until the loss can't decrease anymore.
-
Run play.ipynb
- Run play.ipynb