Deep Mind has a whole bunch of talented and serious people so this is an exciting acquisition.
Deep Mind has a whole bunch of talented and serious people so this is an exciting acquisition.
The only real "deep" learning here is that they used a GPU library and stochastic gradient descent to perform Q-learning updates on a network with 3 large hidden layers. It was an interesting application paper, but I suspect that the Google acquisition is for something more novel than this work.
I hazard it's not very impressed with your Space Invader score, either.
So, I was not impressed by their results on Space Invaders.
Overall, we struggled to learn long-term strategies (finding pure reactive strategies is easy) and to learn to avoid bullets. They did too: "The games Q*bert, Seaquest, Space Invaders, on which we are far from human performance, are more challenging because they require the network to find a strategy that extends over long time scales."
=> that's the real challenge...
[1] http://workinstartups.com/job-board/jobs-at/deepmind-technol...