For example the type of neural net chosen, the reinforcement and other parameters, the rules of the game chosen by the creator are all constraints which are not randomly chosen, so randomly choosing the starting state is not necessarily going to remove preconceptions built in to the model, and it is still choosing some starting state, just not an arbitrary one chosen by the creator.
I'm not sure it really applies to this article though, which is about a different problem - saving the time spent worrying about unimportant decisions or those where you don't have enough inputs and choosing to take a random path.
(A) There are always "preconceptions" present, even if the programmer isn't working on a level of abstraction that lets them consciously realize it.
(B) Programming a system to be too naive can backfire.