5.4 Extra high >> Opus 4.6
5.4 Extra high >> Opus 4.6
I find that for human in the loop Gemini beats both.
It doesn't help that Google offers a bunch of confusing plans in multiple places. I ended up just pasting all their AI plan URLs, at least that I could find, into Claude so I didn't have to figure it out.
https://knowledge.workspace.google.com/admin/getting-started...
They have multiple offerings, they probably will kill some of them very soon. There is no reason to waste your time and money on Google.
I've finished (as in: it's done, it works, and I may never need to change it again) entire projects with ChatGPT and Codex. Sometimes it takes a lot of hand-holding to get there, but it does get there and (with the exception of 4o) it's been improving since the beginning.
In contrast: I can't even get Gemini Pro to give me any answers to the most primitive questions that aren't caked in prima facie lies without at least 4 interactions, in any context, ever. The output is consistently and ridiculously garish with its insessant self-contradictions. It seems to be impossible to actually get anywhere with it.
What am I doing wrong here?