Show HN: Liar's Dice AI from reinforcement learning
dudo.ai
dudo.ai
The code is on https://github.com/thomasahle/liars-dice . I will try to write a blog post about how it works later.
It starts learning from completely random play, but after about a million games the model is close to the Nash Equilibrium. The pytorch model is converted to ONNX runs entirely in the browser.
I also found the CFR papers really confusing for a long time. The papers have a lot of strange notation. But I think I finally got it, and it's very simple! I can write more about it, if I end up writing a blog post about this project.
https://i.postimg.cc/85TPBKgC/Screenshot-20211224-153956.png
EDIT: Also had a game with only 3 dice?
For the second part, if you win a game, you lose a die. The goal is to get to zero!