It’s never useful information. It’s like asking a class of highschoolers to explain in a report why they made certain choices in an art project. The real reason is “i liked it like that” but if you ask them to produce 10 pages of fluff justifying the choices based on nothing they’ll be happy to. (except of course for that one uber diligent kid who actually thought about concepts etc before starting, sorry if that was you, my point is that the AI isn’t like that)
Not that I don't try, but I still get bored halfway through because it sounds like I've read the same thing before.
A more cool question is whether the model is carrying a latent representation of the destination that isn’t yet reflected in the immediate token probabilities?
I don’t know. Anthropic is investigating: https://www.anthropic.com/research/tracing-thoughts-language...
Tested a passage on Pangram and 100% AI generated btw https://www.pangram.com/history/f6b20d8e-4eb4-4bc2-b325-ad39...