HNHacker News
TopNewBestAskShowJobs

zaj00l

11 karma · joined August 14, 2026

submissionscomments
zaj00l··on Run Qwen3.8 27B locally: real numbers from my Mac Studio
I wouldn't say worthless but once you get used to 100+ tokens / s, it's really a visible slowdown, especially when running multiple agents acting on something more than basic prompt processing.

Still, as you said - for day to day, 50-60 tokens / s is a good baseline.

zaj00l··on GLM-5.3-Flash
I get all that.

Then alternatives are:

- Grok - where I absolutely have 0 trust in X.ai's interst in "pushing humanity forward".

- OpenAI and Anthropic - which seem to try to be building the biggest moat they can by pushing to ban open models. And at the same time want to be an Arbiter of what level of intelligence I can use.

- Google and Meta - I don't need to talk about the practices of these companies.

Yes, the terms of service aren't great. But the alternatives aren't great either. I don't believe that a future which OpenAI and Anthropic are pushing for has my best interest in mind.