ACM Prize in Computing Awarded to AlphaGo Developer
acm.org
acm.org
If y'all haven't already, there's a new AlphaGo documentary made by DeepMind on YT: https://www.youtube.com/watch?v=WXuK6gekU1Y. Brought tears to my eyes. Both a triumph for humanity in building an unbelievable machine like this, and a loss for humanity in that the infinite mystery of Go will be diminished, and again a triumph for humanity in Lee Sedol's brilliant win in Game 4...
> What were you thinking when you made that play [move 78]?
> Lee Sedol: Move 78 was the only move I could see. There was no other placement. It was the only option for me, so I put it there.
I do think it is a great achievement how far AI or ML has come. Not to take anything away from the team's accomplishments.
I guess the other members of the team are being compensated well enough at DeepMind that $250k would be more icing than cake, but it still feels weird to see that Silver is the only person named in the article when a number of other world class researchers worked with him on this problem.
For those who aren't that well versed in RL, I recommend watching his lectures at UCL (https://www.youtube.com/watch?v=2pWv7GOvuf0). Really clear explanations that went hand in hand when I was reading Sutton and Barto's introductory book.
I'd just run games, look at results, and endlessly tweak the strategy. Recently I learned how neural networks worked, and realize I could finally make a computer strategy that was competent. It could be trained by playing zillions of games against itself.
My only defense is that training a neural network was impractical on the machines Empire was developed on.
It's hard to resist going back to Empire and doing this.
EDIT: sorry, I found out that source code is available, I will try to find it
EDIT2: Looks like it's up for sale, not open
It's the same algorithm. Mainly a bunch of ad-hoc heuristics.
[0] - https://www.youtube.com/playlist?list=PLqYmG7hTraZDM-OYHWgPe...