There's also plenty of evidence already that RL isn't really the right way to tackle the credit problem. E.g random search is only 10x slower.
Not true. RL has been applied often in production system inside the bigger Internet companies. They are just not published.
and other control problems: https://vmayoral.github.io/robots,/ai,/deep/learning,/rl,/re...
I think not so long ago supervised learning felt similarly 'toy', and now they are the state of the art model for machine translation and object detection. Just because they aren't useful today doesn't mean that they won't be useful.
So you remark means that more education about RL is needed, and OpenAI, alongside other institutions like FastAI or Startcrowd, is helping for this effort.