Reinforcement learning with unsupervised auxiliary tasks
deepmind.com
deepmind.com
Very impressive. I guess the human limit has to do with humans being limited about number of things to track at once? I wonder if they can apply this to optimizing the traffic lights in a big city.