I'm sorry, what?
I'll put it to you again, you sound like me after playing around with gpt 3.5. It was interesting, but it wasn't good. There were still questions as to whether scaling would persist. My doubts were pretty much obliterated. LLMs haven't been mere text completion for a while now. Heck, even a model small enough to fit on my 3090 is now shockingly competent and useful. Does claude sometimes go off on weird tangents making me reel it in? Yeah, but it's genuinely great in present form, and it's the worst it'll ever be.
You're welcome to your opinion. It used to be mine too, but I disagree now. I find pretty much every AI benchmark has shown forward progress, some with non public datasets, and some things have gotten scarily better.