58 karma · joined June 7, 2026
in any case, I've been using open and closed models since sonnet 4, i remember when the best I could get was qwen 3 480b coder, you can definitely feel the gap closing going from that and GLM 4.5, to GLM 5.3, DeepSeek Flash V4.1, Kimi K3 etc, it's reached the point where i wish I had V4.1 at work, it's faster and bullshits me less when I use it in my personal projects. And I have unlimited access to fable 5.1
nothing on the fine print tells you what the weights are, you're just getting Fable 5, whatever that is
Opus 5 came out with better benchmark results than Fable, but it really did not feel better to use at all.
have him start with an overall design doc if his change is 15k, it's definitely worth a design doc.
and then have his contributions reviewed in pieces of 200-300 LoC PRs.
any other solution is trading stability and system knowledge, that's 15k LoC no one is truly familiar with, even if you do try to review it
I'll just cancel my rental agreement, move my furniture across continents, cancel my other subscriptions, oh wait that's got a few month notice period so I'll have to pay up for months I don't need anymore, let alone all the friendships and social connections I made that I never got to say a proper goodbye to.
This is extremely unappealing to skilled workers, and the loss of trust will extend past this current administration, if all it takes is one idiot in office to upend your life, you can't trust that country and its people anymore.
one company deciding a feature freeze is not disapproving anything, in fact, it means things have gotten so bad, apple who used to have higher standards for their releases, had to pull the brakes and fix their shit
i don't think a lot of people know this, but a cluster of GPUs can serve multiple clients without much of a drop in performance, e.i. worst case scenario you band together with 6-16 people to run a 2-3 H100 server to host deepseek V4 Flash or 4-6 to run Pro, and you're getting the same performance as if you ran it alone, this means a lot of companies can afford throwing 50-100k into their own LLM server cluster.
We're at a price point where if you push it further people will move, there's no real vendor lock in, your agent config, skills, MCP servers etc are all reusable with other models and harnesses, so unless you get all providers to collude on a price hike, you risk an exodus of customers