The LLM didn’t generate the next word. Hacker News commenters did. You can see the source of the comment on the results screen.
I propose you do the same things, but only include HN content from before the existence of LLMs. That should ensure there is no bias towards any of the models.
you keep saying LLM when you mean chatbot, i’m not sure if you’re really reading my posts