Show HN: DeepSeek-V4 Latent Reasoning – moving "thinking" into latent space
blog.n.ichol.ai
blog.n.ichol.ai
Therefore output tokens are decodeable, but are trained compressed. So they are approximations of faster thinking.
Interpretability is a mixed bag even with trained tools on top of existing models.
I wanted a knob to keep thinking going until it was sure it was done.