Counter argument: does anything else work this way? E.g. Moores law had an end too right? I would argue that the core tech breakthrough (Transformer-based LLM) has been improved, but no fundamental further innovation seems to have been made. The current architecture fundamentally hallucinates, even Fabel even on trivial problems. I.e. as number tokens increase error likelihood goes to infinity. How then, can this scale recursively to infinity?