HNHacker News
TopNewBestAskShowJobs

lostmsu

6,610 karma · joined January 17, 2014

Threw hundreds of baby transformers into water because their bits-per-byte were too high.

Some fun stuff:

https://borgcloud.org/speech-to-text at $0.06/h

Roxy: iOS hands-free voice AI: https://itunes.apple.com/app/id6737482921?mt=8

Turing Test Battle Royale: https://trashtalk.borg.games

meet.hn/city/43.6534817,-79.3839347/Toronto

Socials: - linkedin.com/in/victor-msu - reddit.com/user/lostmsu - github.com/lostmsu

Interests: AI/ML, Gaming, Networking, Programming, Research, Science, Startups, Technology

---

submissionscomments
lostmsu··on OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
Performance difference it large by all benchmarks. DeepSeek fell behind. It's Kimi K3 or GLM-5.3 now.
lostmsu··on Digital Immortality
Yes, I don't believe in your "spirit", but I certainly believe in mine https://dictionary.cambridge.org/dictionary/english/spirit

I even hinted at the meaning by using "spirits". Next time you may want to use AI to explain usage of words based on their context before jumping to telling me what to do.

I mean you don't even understand that spirit death in the context is not synonymous to bodily death, but a mere consequence of it. Really bad.

lostmsu··on A week of using Codex more than Claude
DeepSeek is much dumber at the moment. It's barely better than Qwen3.8 27B that you can run locally.
lostmsu··on Nearly 25% of U.S. workers are functionally unemployed, economic analysis finds
They are trying to present a new unemployment metric as more reasonable, but throughout the article only give 2 data points making it impossible to judge except from prior bias.
lostmsu··on Show HN: terminal-code – VS Code inside the terminal
I got excited that it's like TurboPascal, but this is basically VS Code over RDP (uses video streaming over kitty protocol).
lostmsu··on Digital Immortality
No, when people die and rot in the ground - that's a complete spiritual death. I assure you my spirits are in a good shape. And I hope the model will too.
lostmsu··on Digital Immortality
> And if that isn’t a comfort, I don’t know what is.

Same, but actually much of your entire life recorded and trained on.

lostmsu··on Unsloth Dynamic 3.0 GGUFs
KLD isn't how that works either. The truth is in the middle and they aren't showing it.
lostmsu··on Unsloth Dynamic 3.0 GGUFs
I would say they do compound until proven otherwise.

Having "Wait, bar is not true, so that won't work" is not necessarily a correction. In fact, the problem is: across a long text it is a correction of a single mistake, but we are talking about thousands here.

But yes, of course that was a rough estimate. But the problem is - we don't really know what we are measuring here. Maybe there's a 2,000,000x difference of intelligence between coding indexes 52 and 50. By some measure that just feels small because that's how we process it akin to audio db.

Regardless the point is KLD and whatever they came up with is not meaningful. And they did not publish comparisons on real benchmarks.

lostmsu··on Unsloth Dynamic 3.0 GGUFs
Cool. Now run TerminalHard and compare to unquantized 27B.

KLD of 1%, or similar error metric that multiplies, on 10000 tokens would give accumulated error of 2,000,000%

lostmsu··on One Oakland police officer made $490k in overtime
It sounds way worse than just napping.
lostmsu··on Cerebras CS-4
KV caching status?

What's the point of 1000tok/s if you have to do prefill on every agentic turn which at 100k depth would make it 1.5 min latency every turn?

lostmsu··on The Benchmarkpocalypse
I wouldn't generalize this to all LLMs. So far I only saw Anthropic ones affected.
lostmsu··on Fixing a Bricked Framework Laptop
My old crappy HP laptop boots with a battery pulled out.
lostmsu··on The Amazon Tax
The ads make sense in capitalism. Without the described effect products would never change. Payment for ads ensures discovery works in adversarial free market.

The problem is that the pay for that goes to Amazon instead of your pockets. I think Brave was actually on to something.

lostmsu··on The Origin of Consciousness (2008)
Why only as oneself experiences? I can easily tell if a horse is conscious or not. The whole matter is rather simple IMO.

Let's say it's variability (number of distinct if you prefer discrete) of stable responses to external stimuli.

lostmsu··on We Tracked a Shipment of Rare Books. It Ended at an Amazon AI Training Facility
If that would be the case, how come only 3 copies are left?
lostmsu··on Semaglutide linked to lower predicted dementia risk
Do you think the lower energy levels also affect your cognition?
lostmsu··on Qwen 3.8 27B is excellent, but it defaults to overthinking things
Glimmer is stupider than 3.6 27B. You can't compare its speed to 3.8 and be done.
lostmsu··on What happens when an LLM never sees material beyond fifth grade?
No, it would have been better to say "I don't know"
lostmsu··on Qwen 3.8 27B
> proxy for totally broken or not

You can't really know that either.

lostmsu··on When Genius Fails: The Intellectual Arrogance of the AI Labs
Your comment is no less ridiculous than his prediction. It is not even 2027 yet. You can't possibly know what AI will look like in 2027.
lostmsu··on Qwen 3.8 27B
> The fact that Unsloth only just started publishing KL divergences shows how unserious the quantization space is.

Just wanted to say that this is a very important point that I totally agree with. People are obsessed with KL divergence, but it is yet to be demonstrated to be a descent proxy for agentic coding benchmarks.

lostmsu··on Accelerating GPT-5.6 Sol Ultrafast
Still no KV caching?
lostmsu··on Grok 4.6
Wow, OpenAI is now 4th after Opus 5, K3, and Grok
lostmsu··on Nvidia Nemotron 3.5 Lightning
> nvfp4

Still trying to lock in, huh?

lostmsu··on LFM2.5 2.6B model competitive with 4x larger models
It's not even competitive with 2x sized Qwen 4B.

Why is Qwen3.5 2B not in the table?

lostmsu··on Launch HN: Stoa Markets (YC S26) – A Marketplace for GPUs and AI Servers
No mention on the website as far as I can see.
lostmsu··on Launch HN: Stoa Markets (YC S26) – A Marketplace for GPUs and AI Servers
No AMD?
lostmsu··on Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
> sadly no breakdown

That's exactly the point. We know short context knowledge stuff does not regress with quantization. But I expect agentic intelligence to suffer greatly.

If I were to pick one bench, I would like to compare quants on TerminalBench Hard. But then Glimmer already loses to 3.6 27B on it by a large margin.

← PreviousPage 3 of 34Next →