As an example, 2026 GPT doesn't even agree with its 2025 self. Last year I asked it to make a hardware comparison and it correctly identified the objectively better option. Recently I asked again and this time it got everything completely backwards.
Gemini's answer was very opinionated and factually correct, whereas Claude gave a more nuanced answer, which was also very good.