Yes, it should be "positive%". That's an interesting LLM-level optimization.
Doesn't this token sampling optimization require using a locally-running model like Llama?
I am presuming that OpenAI doesn't provide direct access to token probabilities in its API.