I got into the habit of running every query I do through all of the SOTA models so I can directly compare the results.
This was originally GPT-4, Claude 3 Opus, and Gemini Advanced. I recently added Meta AI when they launched.
Right now I've sent 486 queries through the first three systems.
The clearest pattern to emerge is that Gemini is terrible, not on par with the other two. There hasn't been a single query that it was the only model who did well. Around 1/4 of the time it gives a clearly inferior answer to the others.
But between GPT-4 and Claude it's less clear. 31 of the 486 queries Claude provided a significantly better answer than the other two but 20 times GPT-4 provided the significantly better answer.
I do think that Claude is a slightly better model but right now it's not a clear enough advantage that I'd recommend it generally. I will say you can probably cancel you Gemini subscription if you're using it though.