6,852 karma · joined June 21, 2024
Also, I feel like the LLMs would have done better if they had started from scratch each time, rather than being burdened by the output from the previous attempt.
That's a non-sequitur.
"Nestle is a great company, considering how much people love their chocolate."
They're both pretty horrible, but I find it difficult to find arguments for why Anthropic is worse than OpenAI, other than their doomtrolling. Which, in the grand scheme of things, doesn't even register.
Edit: forgot about the SpaceX thing.
By raising it from investors.
I'm sure they're doing all kinds of terrible things, like all major companies. I just can't help but like them. Also, this model looks great, and I'll give their subscription a shot next month.
As a society, we decide what we value, and then we pay for it. If you don't like what the group decides, it's your choice to move elsewhere.
So if anyone is wondering if this is just nostalgia, or if the game actually holds up, it's the latter. Just play it.
Everything about this blog post makes me think that I don't want to work with him, not just the "act with urgency" claim. The whole idea that he thought this was a blog post he should write and publish makes me not want to work with him.
I think there is a worthwhile idea in this blog post, which is that you should generally not strongly commit to any specific solution to a problem because you will learn new information while working on the solution, and that it is fine to say, "I was wrong; let's take a step back and rethink this."
But if you were to write a useful post about this, you'd focus on how to decide when to change your approach and when to stick to it, because always rethinking your approach can lead to infinite churn without any releases.
> I think you know better than to accuse me of hurting people
No, I actually really don't. You saw something and posted something that was clearly meant to hurt the person who created it, with absolutely no upside for anyone involved.
If this performs similarly in the real world, we're approaching a level of capability where for most devs, it only makes sense to pay for Anthropic or OpenAI subscriptions if they are heavily subsidized and actually cheaper than these alternative options.
* Oddly, because I perceived Devin as being kind of a joke before trying SWE-2.
Maybe it's a coincidence that the company doing this also tends to have the best models (and other factors certainly play a strong role). But I think it's plausible that focusing on "model welfare" actually makes models better at their tasks.
It's interesting to me that one can look back at the effects that decision has had on the US and say it "was 100% correct."
It's a bit like sitting in the burning ruins of Rome and contemplating that Nero was 100% correct to focus on his music. I mean, I'm glad he got to do what he loves, but maybe 100% is just a tiny bit of an overstatement.
I wonder how much that extends to using LLMs for programming. I assume most knowledge of programming language syntax still comes from training data.
This is false.
> There are many accounts of this on HN itself.
Yes, I saw one just recently where the poster made obviously false statements about starting a company in Germany.
(Before you downvote this comment, please note that I have started multiple companies in the EU. I'm speaking from personal experience, not from something I've read on Hacker News.)