My guess is that in the near* future, reasoning will no longer happen in a way that can be neatly decoded as human language.
*near meaning single digit years, which is far for AI I guess
*near meaning single digit years, which is far for AI I guess
it doesn't seem necessary to read a full CoT exchange. rather a final graph of why a decision was made would be ideal for my usage.
"Latent reasoning" is rather trivial - you can just replace unembed-embed step with a MLP. But labs don't do that largely because they want to read the output of unembed.
Some things are entirely outside of language. Language usually works fine only because most words are encodings of thought patterns that are already present in both parties.
Does an LLM know what blue is? A multimodal LLM probably does, because it has encoders for non-language tokens!