HNHacker News
TopNewBestAskShowJobs

InsideOutSanta

6,730 karma · joined June 21, 2024

submissionscomments
InsideOutSanta··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
Same experience. I often see people say how little they spend on DeepSeek v4.1 flash, but when I put 60 bucks into my account, it was gone in a few days of non-exclusive use. I'm actually curious what the difference is. I used it through pi and opencode, but the harness seemed to have no obvious impact on usage.
InsideOutSanta··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
Yeah, Fable is essentially unusable, it just burns through quota, but Opus 5.5 is great. The $200 plan goes a long way.
InsideOutSanta··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
Grandfathered for a whole month. You're not missing much.
InsideOutSanta··on ChatGPT Pro 500
...while halving the limits on the $200 plan.
InsideOutSanta··on ChatGPT Pro 500
Yeah, with Opus 5.5, my impression is that the OpenAI $200 plan was already worse than the Anthropic $200 plan. Sol 6.1 might have tilted the balance in OpenAI's favor again, but these changes more than negate that.

I'll cancel my OpenAI plan after this billing cycle. Let's see how this plays out.

InsideOutSanta··on Coding Is Not Solved
> If your argument is that LLMs and humans can both make mistakes

It's not, I'm just pointing out that LLMs won't make that mistake.

You could ask an LLM what 1+1 is, and the number of times it says "3" is so small that it makes no sense to worry about it. It will phrase the response differently each time; that's the nondeterminism. But it won't say "3".

> then major question here is why are we building out huge amounts of infrastructure at unsustainable spending levels to enable LLMs to make the same mistakes as humans.

Yes, if we ignore everything else, that seems like a reasonable question. But let's not ignore everything else, like the fact that LLMs are much more productive than humans and likely already make fewer mistakes than the average programmer.

InsideOutSanta··on Coding is not solved
Neither will LLMs. That's not how their nondeterminism works.
InsideOutSanta··on Coding Is Not Solved
> Strange that the world worked before 2024

It must have been a huge shock when you were suddenly transported from a working parallel universe into ours back in 2024.

InsideOutSanta··on On caring for user data: NeoVim caused Vim undo files to be deleted
What do you mean by "obligation"? Nobody is claiming they have some kind of legal obligation. Nobody is saying anything analogous to your "health department" analogy. I think you're arguing against something people by and large are not claiming.
InsideOutSanta··on On caring for user data: NeoVim caused Vim undo files to be deleted
The original analogy was not about selling food:

> If I give away food I have a duty not to poison you. It doesn’t matter you didn’t pay for it.

If I invite you to eat at my house, there is no health inspector and no government. I'm just giving you free food. Would you agree that I have a duty not to poison you in that case?

InsideOutSanta··on On caring for user data: NeoVim caused Vim undo files to be deleted
I published my comment for free, and yet you are criticizing it.
InsideOutSanta··on On caring for user data: NeoVim caused Vim undo files to be deleted
"Your analogy is false because the two things you are comparing are not literally the same, dear sir!" is such a classic Internet discussion trope that it must have a name. If it doesn't, can I please name it? Maybe Perfect Analogy Fallacy?
InsideOutSanta··on On caring for user data: NeoVim caused Vim undo files to be deleted
It's genuinely mind-blowing to me that software can do something obviously bad, someone can point it out, and then someone will link to the license file to say they have the right to do it.

That's such an obvious category mistake that I'm not sure how to respond. It almost feels like a bad-faith interpretation of Wichary's original point.

InsideOutSanta··on On caring for user data: NeoVim caused Vim undo files to be deleted
Ah, yes. The old story of the one-footed man who didn't take the car.
InsideOutSanta··on Analyzing Frontier Model Progress with My Favourite Game: Prince of Persia
PCs were incredibly expensive at the time, though, and barely useful. Most people didn't buy a new one every two years, despite the fast progress in hardware development. So many people, especially kids, were stuck with "ancient" hardware. I had friends who had an 8088 PC or an Apple II in the late 90s.
InsideOutSanta··on Analyzing Frontier Model Progress with My Favourite Game: Prince of Persia
It's so confusing how the actual article is in English, but the prompts are just gibberish. But LLMs are pretty good at deciphering gibberish; I often put our CEO's absolutely atrocious E-Mails into ChatGPT and tell it to explain wtf he wants from me.

Also, I feel like the LLMs would have done better if they had started from scratch each time, rather than being burdened by the output from the previous attempt.

InsideOutSanta··on Toyota is taking the Corolla electric
Perfect.
InsideOutSanta··on DoorDash Spent $1.4M Trying to Stop Mamdani from Becoming Mayor. Now We Know Why
"Simply put, we screwed up. We should have given Cuomo two million so we can keep underpaying Dashers."
InsideOutSanta··on Claude discovers a novel enzyme system with CRISPR-like repeats
This is a bit like saying that shoes are more important than shoe factories. Yes, sure, I can't wear a shoe factory, but we'll all run out of shoes if all the factories are gone.
InsideOutSanta··on What California is learning from solar panels built over irrigation canals
Studies show that the severity of the expected punishment has a relatively small impact on crime (and in some cases, more severe punishments can lead to more severe crimes). The probability of getting caught has a much bigger impact.
InsideOutSanta··on Claude Opus 5.5
People claiming they don't understand or are confused by relatively simple English is a surprisingly common genre of comments on HN. I wonder what causes the people here to have these feelings towards English; my guess is indeed that many people here judge English as if it were a programming language.
InsideOutSanta··on GPT-6 Sol and Luna
> Also, I don't think either of them are horrible. That's honestly a ridiculous take considering how much people in here love their models

That's a non-sequitur.

"Nestle is a great company, considering how much people love their chocolate."

InsideOutSanta··on GPT-6 Sol and Luna
They want to lock people into using the Claude Code ecosystem to make switching to other providers more difficult.
InsideOutSanta··on GPT-6 Sol and Luna
> Also, OpenAI is just a company I'd rather support than Anthropic.

They're both pretty horrible, but I find it difficult to find arguments for why Anthropic is worse than OpenAI, other than their doomtrolling. Which, in the grand scheme of things, doesn't even register.

Edit: forgot about the SpaceX thing.

InsideOutSanta··on GPT-6 Sol and Luna
I think the problem with Anthropic's plan is that Fable just destroys it. If you stick to Opus and below, the $200 plan goes from "using 50% of the weekly quota on the first day" to something much more reasonable.
InsideOutSanta··on GPT-6 Sol and Luna
> I dont know how they make money here

By raising it from investors.

InsideOutSanta··on MiMo v2.6
It's funny, I have the exact opposite reaction. This is probably misguided on my part, but Xiaomi is one of the very few major tech companies that I don't have an immediate strong negative reaction to. Everything I've bought from them, from robot vacuum to mobile phone, has been reasonably well designed, didn't break, and was priced fairly. I also think their car looks badass.

I'm sure they're doing all kinds of terrible things, like all major companies. I just can't help but like them. Also, this model looks great, and I'll give their subscription a shot next month.

InsideOutSanta··on Far-left party wins Berlin election, pledging to nationalise housing
I mean, that's how taxes work. I don't have children, and yet I pay for schools. I barely ever drive anywhere, and yet I pay for roads.

As a society, we decide what we value, and then we pay for it. If you don't like what the group decides, it's your choice to move elsewhere.

InsideOutSanta··on Grim Fandango Puzzle Document (1996) [pdf]
I first played it in my 40s when the remaster came out, and it's one of the best games I've played. It holds up incredibly well, particularly with the modernized controls.

So if anyone is wondering if this is just nostalgia, or if the game actually holds up, it's the latter. Just play it.

InsideOutSanta··on I am often wrong
> Heavens. Is it possible you're reading into it a bit? He's miserable to work with? From this one blog post?

Everything about this blog post makes me think that I don't want to work with him, not just the "act with urgency" claim. The whole idea that he thought this was a blog post he should write and publish makes me not want to work with him.

I think there is a worthwhile idea in this blog post, which is that you should generally not strongly commit to any specific solution to a problem because you will learn new information while working on the solution, and that it is fine to say, "I was wrong; let's take a step back and rethink this."

But if you were to write a useful post about this, you'd focus on how to decide when to change your approach and when to stick to it, because always rethinking your approach can lead to infinite churn without any releases.

Page 1 of 34Next →