Using Keras and Deep Deterministic Policy Gradient to play TORCS
yanpanlau.github.io
yanpanlau.github.io
[1] http://vizdoom.cs.put.edu.pl/ [2] https://deepmind.com/research/dqn/
Anyone interested in this type of research: consider cloning the repo and implementing this modification, it would make a great starter project.
Given enough training, would the car learn to find the apexes of turns?
It is probably an easy local maxima with relatively fast convergence
Far harder are
a. "planning" - finding the optimal path through sequential turns
b. generalizing the learned experience to a new, unseen situation
Would love the get feedback from the author on this
Just like in human world : You first learn how to drive before you learn how to drift the car.
How would the existing agent fare on a brand new track ?