869 karma · joined December 7, 2012
So, folks that have actually used this already, what’s it actually like?
This is terrible. I have no desire to read your Claude conversation. I don’t even want to read mine if I do use Claude.
The code is the artifact. There’s no value in trying to make the noisy, verbose conversation one. It’ll likely be dead in future years anyway.
Any/all work!
I wouldn't fully trust LLMs with this :p BUT this project is still really freaking sweet.
It's going to get ugly when large swaths of the human population struggles to get food & clean water.
Probably because anthropic is pouring $40 million dollars into a political pact to regulate models. And they have not been quiet about wanting to ban/regulate OSS models either.
I have no idea why HN still treats them like ~their~ they are the ethical good guy.
Edit: I can’t English. At least we know this came from my dumb swishy brain
The models write good/great code. I'm very happy to never write code again but models are no where near good enough to run off on their own without a human in the loop. Models LOVE to cheat. Just peek at the tests they write.
And before you come at me, I have built 7 products in the past 1.5/2 years with agentic engineering, all with users. One of those projects is _dead_ because I let the vibe go too hard at the same time Anthropic decided to nerf both their harness and their models silently. If you care to look at the source: https://github.com/Robdel12/OrbitDock I spent a week or so and like a billion+ tokens trying to refactor and save it. It just wasn't worth it.
I wish people would be pragmatic about this. I get the dream is to let it do everything and not to be in the loop, because being in the loop is exhausting. But if you want to make whatever you're building be robust and survive more than 6 months, you have to. I don't care how good your tests, plans, skills, etc are. At some point the model will have to decide something and it'll be the wrong one. Compounding the slop from there forward.
This is pretty wild but also I think this is doing a lot of heavy lifting here. This was not a model everyone has access to. I mean, still insane.
So, It’s neat to see something competent! Imagine if they modeled what cutting off the natural draining to the Everglades would do :p
https://www.anthropic.com/research/2028-ai-leadership
They are already starting that now.
There ya go, the rewrite was for marketing.
They’re trying to scale from 1 billion commits last year to over 14 billion this year. I have zero desire to try and manage that scaling. Basically being DDOS’d by agents all day now.
No healthy engineering team is going to do that. And I’d want to distance myself as far as I could from a project that behaves like that.