Also, whatever loose rules it has are more literary than mathematical. Plot twists often work.
Also, whatever loose rules it has are more literary than mathematical. Plot twists often work.
No, it is clearly not, and that is a very easily testable hypothesis.
Thank you for sharing.
For a number of years we've been basically showing the first to be the case, especially as the model is scaled and the context increases, differentially against the second. String density probabilities can be surprisingly brittle, to be honest. The curse of dimensionality applies to them too, believe it or not, which I believe is why topic discussion, reasoning, and integration over longer distances of text is that differential test that shows pretty clearly that substring memorization/text density stuff is not 'just' what the model is learning. Because mathematically/statistically/from an information density perspective/etc etc otherwise it would be basically impossible, I think.
That's my best understanding, at least.
In the analogy of the essay, your argument would be like saying that reality cannot be simply the application of quantum physics, because you are allowed to make new rules like Calvinball within reality which are different from the rules of quantum physics.
We know there's no deeper level to the simulation/game because we have the entire "game history" (the chat history) and we understand it in approximately same way that the LLM does. (That's what the LLM was trained to do, understand and respond to text the same way we do.) We know that the bot has no hidden state when it's not the bot's turn because of how the bot's API works.
So there's nowhere for a deeper simulation to live. It's as shallow as it looks.
More:
https://skybrian.substack.com/p/ai-chats-are-turn-based-game...