Based on the discussions here it seems that every model is either about to be great or was great in the past but now is not. Sucks for those of us who are stuck in the now, though.
Based on the discussions here it seems that every model is either about to be great or was great in the past but now is not. Sucks for those of us who are stuck in the now, though.
https://status.anthropic.com/incidents/72f99lh1cj2c
Suggesting people are "out of their mind" is not really appropriate on this forum, especially so in this circumstance.
This most definitely feels like people analyzing the output of a random process - at this point I am feeling like I'm losing my mind.
(As for the phrasing I was quoting the OP, who I believe took it in the spirit in which it was meant)
[1] https://news.ycombinator.com/item?id=45183587
[2] https://news.ycombinator.com/item?id=45182714
> New features like this feel pointless when the underlying model is becoming unusable.
I recognize I could have been clearer.
And for what it's worth, yes, your comment's phrasing didn't bother me at all.
They were wrong, but not inappropriate. They re-used the "out of their mind" phrase from the parent comment to cheekily refer to the possibility of a cognitive bias.
Yes, but I'll revisit.
On that note, I strongly recommend qwen3:4b. It is _bonkers_ how good it is, especially considering how relatively tiny it is.
FWIW, Codex-CLI w/ ChatGPT5 medium is great right now. Objectively accelerating me. Not a coding god like some posters would have it, but overall freeing up time for me. Observably.
Assuming I haven't had since-cured delusions, the same was true for Claude Code, but isn't any more.
Concrete supporting evidence: From time to time, I have coding CLIs port older projects of varying (but small-ish) sizes from JS to TS. Claude Code used to do well on that. Repeatedly. I did another test last Sunday, and it dug a momentous hole for itself that even liberal sprinkling of 'as unknown' everywhere couldn't solve. Codex managed both the ab-initio port and was able to undig from CC's massive hole abandoned mid-port.
So I'd say the evidence points somewhat against random process, given repeated testing shows clear signal both of past capability and of recent loss of capability.
The idea that it's a "random" process is misguided.
You mean like our human brains and our entire bodies? We are the result of random processes.
>Sucks for those of us who are stuck in the now, though
I don't know what you are doing- but GPT5 is incredible. I literally spent 3 hours last night going back and forth on a project where I loaded some files for a somewhat complicated and tedious conversion between two data formats. And I was able to keep going back and forth and making the improvements incrementally and have AI do 90% of the actual tedious work.
To me it's incredible people don't seem to understand the CURRENT value. It has literally replaced a junior developer for me. I am 100% better off working with AI for all these tedious tasks than passing them off to someone off. We can argue all day if that's good for the world (it's not) but in terms of the current state of AI- it's already incredible.
It might not be a junior dev tool. Senior devs are using AI quite differently to magnify themselves not help them manage juniors with developing ceilings.