Presumably the argument is that training a neural net from a basis of complete ignorance is inefficient because we have facts with which we can initialize the model.
So far as applicability to TFA, we can and probably should initialize or bias models that select candidates so their inferences reflect our values.