The reasoning is great in opus, unbeatable at the moment.
I understand what you mean, it becomes disappointing on more niche or specific work. It’s honestly a good thing to see these models are not really intelligent yet.
I use it for reviewing existing code, specifically for a components-based framework for Godot/GDScript at [0]. You can view the AGENTS.md and see that it's a relatively simple enough project: Just for 2D games and fairly modular so the AI can look at each file/class individually and have to cross-reference maybe 1-3 dependencies/dependents at most at any time during a single pass.
I've been using Codex, and it's helped me catch a lot of bugs that would have taken a long time on my own to even notice at all. Most of my productivity and the commits from the past couple months are thanks to that.
Claude on the other hand, oh man… It just wastes my time. It's had way more gaffes than Codex, on the exact same code and prompts.
I still get really mad at AI sometimes and I am not sure whether I could use AI for coding full time.
(Codex broke my git a few days ago.)
I use Codex regularly and Claude is shit in comparison, from its constant "Oops you're right!!" backtracking to its crap Electron app (if their AI is so good why can't they make a fucking native app for each OS?)
Hell right freakin now I asked it to implement something and got a weird "Something went wrong" API error
Maybe you're too easily frustrated. Or your existing code reads like your comments.
I haven't had any such frustrations with Codex
Claude is specially annoying because of their submarining and people thinking it's the best
and other comments further back in my history
> none of their issues warrant a tantrum on a public forum
I don't get frustrated if a problem is genuinely difficult to solve and the product creator is trying their best,
I get frustrated when a problem has been solved by other similar products but a specific creator or provider refuses to follow suit and fix their shit.
Claude's Electron app vs. Codex's native app is one such example right off the first impression of both products.
Codex definitely "feels" more native than Claude: Proper menus etc, like when you right-click on a session in the sidebar, Codex shows an actual context menu, whereas Claude reveals its HTML-rendered jank and highlights the word you right-clicked on as if the sidebar item is just a plain textbox, ugh
"Why"
"HAVE YOU SEEN THEIR CONTEXT MENU"
Claude's AI. itself. is. trash.
Their UI/UX makes it worse.
The constant fanboying/paid PR around it makes it the worst.
Especially since you’re factually wrong and you’re bad at arguing when your temper gets the best of you.
Just right NOW it generated GDScript code with SPACES instead of TABS for indentation despite me explicitly asking for tabs and everything else in the project using tabs.
Your use case is either extremely simple, involving extremely common scenarios, or you're too lazy to verify Claude's output.
Either way, it's not worth the hassle to convince someone using a bad product that it's bad; nothing's gonna improve, and you're the one being harmed by stubbornly gaslighting yourself and others that it's good. The rest of us will escape by never subscribing again.
I've already wasted too much time on this thread. I keep saying I am using BOTH Claude AND Codex SIDE BY SIDE but that's somehow getting filtered out by your fanboy glasses. I could post screenshots of the outputs of both agents for the SAME prompts but you would still not understand.