They’re undoubtedly getting much smarter overall, but also much weirder. Before they were just trying to model our behavior, only really having us to learn from.
Now they’re literally spending thousands of years writing bash scripts in some kind of Sisyphean dreamscape, talking to each other about Goblins and Seams and smoke tests, and coming back as idiot savants.
I don’t even try to police how Claude talks or works anymore. Best practice used to be to nudge them towards whatever part of the distribution of behavior you think they should exhibit in a particular situation, because they were role-playing what a human in a particular situation would do, and if you didn’t tell them how to do it they’d just role play something worse. Now the inclination to do things the way they learned it in Agent University is so strong, they’ll literally spend more tokens re-assuring themselves and you that they are Doing It Your Way, and reminding themselves not to do give in to temptation, than you could ever prompt out of them. They’re going to spend your money thinking about goblins anyway so just let them
I think the tradeoff to Claude being so needy is that if you let models just run away with an inaccurate or incomplete understanding of what to do, they can go really far off the rails AND spend a lot of time/money doing it AND come back with something that literally doesn’t make sense or doesn’t work.
I prefer dealing with Claude’s reliable cringe to the aloof model that tries to play it cool when it needs help.
:!pkill -KILL vim