(I have this weird feeling that you're not upset with blue/green deployments...)
2,861 karma · joined September 20, 2022
(I have this weird feeling that you're not upset with blue/green deployments...)
A common tactic is to used a big brain model like Opus for planning and reviewing, and a cheaper model for execution.
Non-pedantic answer: I totally agree with you. Opus 5.5 is totally knocking it out of the park IMO.
So what is the value prop then? Just basic supply and demand?
FWIW I have definitely noticed OpenAI's emphasis on efficiency and value in the last year, so that part isn't new to me... I just thought there was more to it then that.
On the other side of the fence, the aversion to using the tool seems equally cultish, like not working on the Sabbath.
If it makes you feel any better, I don't think this tech is "AGI" and that it seems to have some fundamental limitations that make it perform worse than a 3-year-old can perform on many tasks. But it's foolish to avoid something that could cause you to be replaced by some idiot that can do untold damage with just a few prompts, all because you didn't (assuming you work in tech) want to stay on top of the tools available to you.
Also, you are now using the tool that is best suited to "use AI". You will hit those rough patches, and it's good to know how to work through them. But sometimes there is just no time for such bullshit, and you now have a tool at your disposal that can just sweep these problems away. It is not without its downsides, but this is a clear "dumping the baby out with the bathwater" scenario.
Your post has the same energy of the old codgers who won't "learn computers" and impotently swim against the tide. It's fine if you want to do that (you do you), but you are swimming in the wrong direction IMO.
(If you don't work in tech, the magnitude of my recommendations would soften towards you in particular, but the velocity would remain the same.)
Have you ruled out the possibility that your system prompt, AGENTS.md, or increasing codebase complexity are not to blame?
Half the price when it launched, or after the price dropped by 75%?
I assume it's a subsidy to get more training data.
EDIT: Okay downvoters, what's your take on why they're giving away Luna for so cheap?
The internet and its consequences have been a disaster for the personal computer user.
However, I believe that runtime model quantization is possible with some publicly-available inference engines (e.g. vLLM), so its not beyond belief that the closed labs do quantize at runtime, either to allocate compute, or to nudge users towards a preferred model (e.g. make the incumbent model dumber to push people to use the latest-and-greatest model, or vice versa to ease the load on the latest model, which is typically larger than the old one).
EDIT: I forgot (and am shocked) that HN still doesn't seem to support Markdown-style links.
Yes, and now that they've achieved market saturation, it's time to pull up the ladder.
Thanks for all the free work, suckers!
I'll just use my one-size-fits-all AGENTS.md file and tweak it when the one of the clankers screw up. I don't have time for such busywork.
Actually, I will append extra rules to CLAUDE.md (which imports AGENTS.md) since there is a hook there, and Claude has its own foibles. So I'll backpedal a bit there.