"Statistical token generators" doesn't even count as stating the mechanism. If I write a Perl script that randomly chooses between "cat", "dog", "ape" tokens, is that an LLM? What if I train it by feeding it a library of books where it tracks the statistical frequency of each of these and then emits them? Where's my trillion.