Over the past year there have been advances in making models smaller while keeping performance high.
So if that continues then he is wrong unless he is defining LLMs in a strict way that does not include new improvement in the future
So if that continues then he is wrong unless he is defining LLMs in a strict way that does not include new improvement in the future
Humans are able to begin to generalize with a single persons experiences over less than a year, so the fact that LLMs cannot with billions of person-years of information could be an indicator of their inability to generalize no matter how much training data you throw at it.