My personal experience indicates this, AI enhances me but cannot replace me
Been doing something closer to pair programming to see what "vibe" coding is all about (they are not up to being left unattended)
See recent commits to this repo
- Claude Code is a Slot Machine https://news.ycombinator.com/item?id=44702046
- GPTs and Feeling Left Behind: https://news.ycombinator.com/item?id=44851214
- I also used the Emperor/clothes metaphor: https://news.ycombinator.com/item?id=44854649
And just so we are clear, in the only current actual study measuring productivity of experienced developers so far, it actually led to 19% decline in productivity. https://metr.org/blog/2025-07-10-early-2025-ai-experienced-o...
So, if the study showed experienced developers had a decline in productivity, and some developers claim gains in theirs, there is high chance that the people reporting the gains are...less experienced developers.
See, some claim that we are not using LLMs right (skills issue on our part) and that's why we are not getting the gains they do, but maybe it's the other way around: they are getting gains from LLMs because they are not experienced developers (skills issue on their part).
I'm an experience (20y) developer and these tools have saved me many hours on a regular basis, easily covering the monthly costs many times over
It's also not the only data, here's one with results in the other direction
https://medium.com/@sahin.samia/can-ai-really-boost-develope...
It’s interesting they showed the biggest gains for junior developers. The other study showing productivity losses for experienced developers. That suggests these tools are a lot more helpful for junior developers compared to senior developers at the moment.
Theore you use it, the more you get a feel for when and when not to use it
Making untested garbage faster to check off tasks quicker. Reopen rate please? New bug task rate? Nobody looked...
> The study also monitored code quality via build success rates. Importantly, increased productivity did not come at the cost of more errors, showing that Copilot helped developers code faster and more accurately.
It builds therefore it works. And the test suite was also AI generated?
Feels like we're back in primary school learning programming.
Classic case of gaming the metrics.
- people don't outsource, they pair-program, because these things cannot do complex tasks on their own
- they are quite good at running tests with coverage, inspecting the results, and fixing both the tests and code
- people make mistakes, expecting AI to be perfect is unreasonable, they are tools, not replacements
> Making untested garbage faster to check off tasks quicker
> Feels like we're back in primary school learning programming.
> Classic case of gaming the metrics.
plain view bias causes others to discount your opinion
This is the key. These tools are an improvement for many people, but others pooh-pooh them for not being perfect. Working in a team with other programmers (or looking at my own older code) I often see mistakes, often obvious to me now.
You forgot to add: first time users, and within their comfort zone. Because it would be completely different result if they were experienced with AI or outside of their usual domain.
You are also misrepresenting the literature. There are many papers about LLMs and productivity. You can find them on Google Scholar and elsewhere.
The evidence is clear that LLMs make people more productive. Your one cherry picked preprint will get included in future review papers if it gets published.
Go check it!
I have very rarely needed the site: modifier, and even knowing this exists sets people apart, and reinforces there is skill to it
You can hammer in nails with your knuckles. Do you want to do that daily? There are people who can, they're still worse at it than a plain hammer or someone even using a right rock. Much less an assembly line.
Consider that the frequency of replies along those lines might be evidence that there's something to it. It's not necessarily true, of course, but if it's false then you need to explain why so many people believe otherwise.
In the end, every tool I tried felt like I was spending a significant amount of time saying “no that won’t work” just to get a piece of code that would build, let alone fit for the task. There was never an instance where it took less time or produced a better solution than just building it myself, with the added bonus that building it myself meant I understood it better.
In addition to that I got into this line of work because I like solving problems. So even if it was as fast and as reliable as me I’ve changed my job from problem solver to manager, which is not a trade I would make.