The point is that current methods are unable to get more than the current state-of-the-art models' degree of intelligence out of training on the totality of human knowledge. Previously, the amount of compute needed to process that much data was a limit, but not anymore.
So now, in order to progress further, we either have to improve the methods, or synthetically generate more training data, or both.