I've tried all the top models. GPT4 beats everything I've tried, including Gemini 1.5- until today.
I use GPT4 daily on a variety of things.
Claude 3 Opus (been using temperature 0.7) is cleaning up. I'm very impressed.
I use GPT4 daily on a variety of things.
Claude 3 Opus (been using temperature 0.7) is cleaning up. I'm very impressed.
I've continued to test. Definitely wouldn't call it a step function, but love that it's genuinely competitive with GPT4, and often beating it.
I am starting to see some cracks-
It's struggling with more hardcore / low-level programming tasks, but dealing well with complexity / nested abstraction with proper prompting.
It sounds much less AI-y when it talks, like better variation / cadence which I think was what sold me so hard at first.
Otherwise your comment is not quite useful or interesting to most readers as there is no data.