Coding agents have crossed a chasm
blog.singleton.io
blog.singleton.io
“Disclaimer: I’m the CEO of a company that sells agents as a service”
at the top of and article promoting said agents.
The experience of long-term software engineers (e.g. antirez) who don’t have a horse in the AI race tends to line up much better with my own.
Also really like this one: https://diwank.space/field-notes-from-shipping-real-code-wit...
If you believe that agents will replace software developers like me in the near term, then you’d think I have a horse in this race.
But I don’t believe that.
My company pays for Cursor and so do I, and I’m using it with all the latest models. For my main job, writing code in a vast codebase with internal frameworks everywhere, it’s reasonably useless.
For much smaller codebases it’s much better, and it’s excellent for greenfield work.
But greenfield work isn’t where most of the money and time is spent.
There’s an assumption the tools will get much better. There are several ways they could be better (e.g. plugging into typecheckers to enable global reasoning about a codebase) but even then they’re not in replacement territory.
I listen to people like Yann LeCun and Demis Hassabis who believe further as-yet-unknown innovations are needed before we can escape a local maxima that we have with LLMs.
My experience matches theirs - Claude Code is absolutely phenomenal, as is the Cursor tab completion model and the new memory feature.
Charlie Marsh seems to have much better luck writing Rust with Claude than I have. Claude has been great for TypeScript changes and build scripts but lousy when it comes to stuff like Rust borrowing
I'll add - they do seem to do better with Go and Typescript (particularly Next and React) and are somewhat good with Python (although you need a clean project structure with nothing magic in it).
--
So I suppose the chasm is that actually doing programming is dead, or quickly dying, and if that's the thing you actually enjoyed doing, then tough luck.
This era sucks. The suits have finally won.
(emphasis mine)
If not, just do it for yourself.
> I don’t even look at the code anymore - I describe what I want to Claude Code, test the result, make some minor tweaks with the AI and if it’s not good enough, I start over with a slightly different initial prompt.
Honestly, does the author and anyone else using this workflow find this way of working enjoyable? To me programming is not entirely about the end goal. It's mostly the small bursts of dopamine whenever I solve a particular problem; whenever I refactor code to make it cleaner, simpler, and easier to read; whenever I write a test and see it pass, knowing that I'm building a safety net to safely refactor in the future. And so on.
Yes, the feeling of accomplishment after shipping a useful piece of software, be that a small script or a larger part of a system, is also great. But the small wins along the way are the things that make me want to keep programming.
This way of working where you don't even look at the code, but describe the system specs in prose, go back and forth with an extremely confident but highly error prone tool, manually test the result, and repeat this until you're satisfied... doesn't sound fun or interesting at all.
AI (and before that, corporations) makes skepticism more and more a basic survival skill.
Since this is partly an experience report, it is only as trustworthy as its author, whoever that is. What is this person risking by writing it?
The content seems plausible to me. However, what I’m missing here is:
- how does he test?
- how does he keep himself sharp while his tools are doing so much?
- How does he model the failure modes of this approach, or does he just feel it?
I am not having the same feeling of success as this guy is as I experiment with the same tech. Maybe he’s better than me at using it. Or maybe he’s easily impressed.
I do still read the code _except_ when I am consciously vibe coding a non production thing where I will know empirically that it worked or not by using it.
I’m definitely not using agents to do all my coding (as I hope is reasonably) clear from the post. But they have crossed this line from pointless to try to genuinely useful for many real world problems in just the last couple of months in my experience.
Not because I’m certain that the jobs won’t be there, but because I think it’s a credible risk. A career is too important to gamble.
It's not true if your humans are on controlled substances all the time, it is true if we are talking about real humans.
I've been testing coding agents on real code and I can say without a doubt that they make worse mistakes than humans.
Snarky and dismissive, sure. But the Wii wasn't a "1 to 1 motion matching" machine no matter how many people insisted it was. It was just "better than anything before had ever been". Which is not the same thing as "good". I'm not holding anything against anyone. The Wii was an incredible console, an LLMs are an incredible technology. I'd just like to read some thoughts on the tech from people who are more aligned with myself in their discernment. If, for nothing else, some variety.