The model is trained on large volume data, correct? Why would it get such a simple fact incorrect?
Also, the knowledge can be kind of siloed. You often have to come at it in weird ways. Also, they are not fact-bases. They are next-token-predictors, with extra stuff on top. So if people on the internet often get the answer wrong, so will the model.
Can't wait to unleash Pluto questions.