There's also plenty of evidence already that RL isn't really the right way to tackle the credit problem. E.g random search is only 10x slower.
There's also plenty of evidence already that RL isn't really the right way to tackle the credit problem. E.g random search is only 10x slower.
and other control problems: https://vmayoral.github.io/robots,/ai,/deep/learning,/rl,/re...
I think not so long ago supervised learning felt similarly 'toy', and now they are the state of the art model for machine translation and object detection. Just because they aren't useful today doesn't mean that they won't be useful.
Not true. RL has been applied often in production system inside the bigger Internet companies. They are just not published.
So you remark means that more education about RL is needed, and OpenAI, alongside other institutions like FastAI or Startcrowd, is helping for this effort.