ReBeL: A general game-playing AI bot that excels at poker and more
ai.facebook.com
ai.facebook.com
Underlying paper here. Originally published several months ago and updated with more information last week.
Of course, this limited the games that we could simulate to purely deterministic games (checkers, chess, go, etc.). Any games that included an aspect of chance required a hack like a "dice player" or a "deck player" that would add the random aspects of the game. Of course, this led to other problems, since the engines would try to calculate the current state of the game based on the "optimal" play of the random player.
This is a much more interesting approach, and I imagine will prove to be far more useful.
There was some discussion of this in the AlphaGo Zero blog post from a while back: https://deepmind.com/blog/article/alphago-zero-starting-scra...
Imagine searching through the entire tree - billions of nodes due to the branching of chess. Now use RL to help decide where to search and branches to prune