> The LLM is not suited to giving deterministic answers to math problems.
Less so with formal mathematics proofs maybe but I think in general humans don’t provide deterministic answers to math problems or questions either. Humans get it wrong all the time and when you ask a human to solve a problem they may solve it in a different way than before.
But even if you have 5x6 memorized it's not deterministic that you answer 30. It's just highly probable.
People don't understand what correlation is and they assume the mapping between a human brain and an LLM is 1:1 in every case where it matters, which is a religious and not scientifically based belief.
The nice thing about math is it can easily plug into a tool, making it even less of a concern.
5.6 on Instant mode can knock out 3 digit multiplication just fine.
TLDR: Astra has 8.6x better odds of doing a reasoning task without CoT than the
next best model (Fable 5.1), and can do 7.2 serial arithmetic steps in a forward
pass vs 4.1 for the next best model (Gemini 3.8 Flash/Fable 5.1)
https://www.lesswrong.com/posts/eRmzz8J8Qkzqvzrgg/astra-can-...