Along this line of thought: was it a massive oversight for them to not train the model to say "math detected, let me pass that to a solver" instead of trying to guess what token should come next in a math problem?
GPT is quite useful, but not because it solves the problem of "I don't know where the question I have is answerable by a calculator"
So personally I think the problem is that people see "this product does X" and interpret that to mean that it does X well. I don't think it's necessarily bad that we're seeing an explosion of AI tools that are a bit underwhelming if people understood it as such -- we're on, after all, a site with a heavy startup focus and saying "your product doesn't do everything that I want" is a bit antithetical to that.
But yeah specifically for this one there are arguments that "X is not even possible, especially not with this approach" so it's a bit more egregious.
(And no one cares that you used to work at Microsoft or whatever).