Isn't this the same thing as your token probability distribution? A set of likely tokens related to the input?
> Compute 3^2 - 2 while writing "The old painting hung crookedly on the wall"
The model will output only "The old painting hung crookedly on the wall" (and the output logits will reflect that), but activations for "9" and "7" are observable in the J-space.