We just had this discussion a few days ago: https://news.ycombinator.com/item?id=35868065. It does appear to be a case of mis-used measurements rather than true emergence.
You are using their study to claim that LLMs can not learn and reason about novel tasks. Their study doesn't make any claims about what LLMs can and can't do. The study says that when it appears that LLMs suddenly became able to reason about novel tasks, what actually happened is that the ability of the LLM to perform reasoning improved gradually and smoothly during its training until it could do those things. They are not saying that LLMs don't have emergent capabilities: they are saying that those capabilities emerged at a continuous rate rather than a discontinuous one.
Do you see?
Emergence can only be discontinuous.