I was using Claude Opus to code before and now I'm back to ChatGPT. GPT-4o is faster, doesn't generate placeholders and works way better for me because of the larger context.
It's also what people think in blind tests: https://arena.lmsys.org
They should really be streaming the content at the same time, based on the slowest responder.
I'm getting the strong impression that 4o is significantly weaker than 4, at least for dealing with coding snippets.
I suspect it could be related to whatever it's using as language detection, because many others don't experience this. It glitches hard on language, often responding in the wrong one.
When I return to 8x7b from gpt-4 it feels like I just shook off an unbearably boring guy and met a normal one, both very similar in knowledge (and unable to perform complex tasks).