Like "Please reply T if you are confident about the text that you are going to generate, and F if you are not".
But if this could be done by architecting the neural networks, they would perform much better.
Like "Please reply T if you are confident about the text that you are going to generate, and F if you are not".
But if this could be done by architecting the neural networks, they would perform much better.
We don't fully understand how they work or what their limits are.
It's also worth remembering that this is not just a markov chain as many people seem to think, it doesn't simply remember what words come next in its training set, because statistically if you take a random 10 consecutive words from a piece of text, chances are its never been written before. (also the trained model is much smaller than the size of the dataset, so to simply remember everything it would have to be the worlds best compression algorithm by orders of magnitude) Thats why we need an AI here, to learn the general rules of language so it can respond to chains of words it has never seen before.
The sense that we "dont understand how they work" is that we dont know what the "rules of language" that it has learned are.
This is no more helpful in understanding AI's than is knowing that human brains operate according to the laws of physics is helpful in understanding the human mind.
[1] https://skybrian.substack.com/p/dont-settle-for-a-superficia...
[2] https://skybrian.substack.com/p/ai-chats-are-turn-based-game...