API doesn't have these issue, or didn't a week ago when I was using it often, but something happened with the web model, which I guess is understandable, as the main reason I was using it vs the API was cost savings.
If they’re doing something like this that would probably explain it.
GPT 4 is now a lot faster and struggles much more. And in theory being faster should be an advantage to an extent, even if the model needs more prodding, as you gain a more seamless back and forth instead of waiting half a minute.
Some people say that people think it's nerfed because when it was slow we associated it with "thinking harder" but I doubt that's the case, and the way to measure that is by loss of ability, i.e. if it is stuck more often as a result of itself and not a bad prompt.
And I guess the trend is that anyone who uses it frequently uses 4 and 3.5 for basic stuff only.
Loads of people claim this, I haven't tested it myself though.