If the robot already knows "how to" the happy path, the training difficulty falls severely at least if it can continue after a recovery.
But yeah, I think a better way to put it is that sampling the happy path would indeed make the failure case easier, but sampling just happy paths is far from sufficient from completing even some of the simplest human tasks with failure.