Show HN: Easily train AlphaZero-like agents on any environment you want
github.com
github.com
You'll also likely want to mention the "needs python >= 3.8" in the readme https://github.com/s-casci/tinyzero/blob/244a263976cd9a09f5f... OT1H, I would hope folks are keeping their pythons current, but OTOH dev environments are gonna dev environment
[1]. https://en.wikipedia.org/wiki/Game_Description_Language
[1] github.com/Entze/pyggp
While much of the course material survives [1], those rulesets do not. The only GDL example I could find was the somewhat trivial example of Tic Tac Toe, see section 2.6 Tic Tac Toe Game Rules here [2].
[1]. http://ggp.stanford.edu/public/lessons.php
[2]. http://ggp.stanford.edu/chapters/chapter_02.html
An email for Michael Genesereth, teacher of the course, is on the course website. I might shoot an e-mail and see if he has GDL files to share.
I haven’t used it in a few years but certainly was the standard back then
> get_legal_actions(): returns a list of legal actions
What's the expectation around your actions? It's not just 0..n for current actions with any arbitrary ordering, right? There needs to be some consistency between steps for training.
(You could also move straight to MuZero variations: https://arxiv.org/abs/2106.04615#deepmind https://openreview.net/forum?id=X6D9bAHhBQ1#deepmind https://openreview.net/forum?id=QnzSSoqmAvB )