Claude and GPT tell me all the time what they are confident and what they're not confident in. This is partially useful, but it's not a guarantee of anything, since they can be confidently wrong or unsure and correct.
Same with Jav.
Arguing that Jav in some way more useful than current LLMs because it includes a confidence score smells of someone who hasn't updated their priors since GPT-4/4o and it's particular issues.
Jav is cool because it's very fast, very cheap, and "good enough". That's it.