HNHacker News
TopNewBestAskShowJobs

lostmsu

6,610 karma · joined January 17, 2014

Threw hundreds of baby transformers into water because their bits-per-byte were too high.

Some fun stuff:

https://borgcloud.org/speech-to-text at $0.06/h

Roxy: iOS hands-free voice AI: https://itunes.apple.com/app/id6737482921?mt=8

Turing Test Battle Royale: https://trashtalk.borg.games

meet.hn/city/43.6534817,-79.3839347/Toronto

Socials: - linkedin.com/in/victor-msu - reddit.com/user/lostmsu - github.com/lostmsu

Interests: AI/ML, Gaming, Networking, Programming, Research, Science, Startups, Technology

---

submissionscomments
lostmsu··on Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
Are you running inference in parallel? 70 tps seems low for parallel execution.
lostmsu··on Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
This release is not a meaningful improvement in any metric over 5 months old Qwen 3.6.

DS v4 Flash update maybe, but it is too big for typical Joe's desktop.

lostmsu··on Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
From my perspective it doesn't make sense to talk about the number of parameters. What matters is model size in bytes and its performance at that certain size.

Meta actually relesed official 4 bit quants in 17GB, but I haven't seen any indication that training was quant-aware, so the quants are not going to have same performance. 3.6 27B has official FP8 quant that AFAIR was trained with quantization awareness.

The best example is last year's gpt-oss which was released prequantized in mxfp4 so 20B parameter model was under 14GB and 120B was under 70GB right away.

lostmsu··on Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
They say it is trained with quantization awareness, so it should only be 15GB or so. Qwen was only trained in FP8 with QAT.

UPD, NVM, got misled by comments here. It is actually almost 60 GB so much larger

lostmsu··on Silicon Valley misreads science fiction and undermines democracy
State of the art of philosophy is IMHO Popper and scientific method.

Everything that came after is BS, everything that came before is obsolete.

lostmsu··on Hetzner Experiments Platform: Inference API
Requires login. Any public page?
lostmsu··on Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
It seems worse than 3.6, but a bit smaller.

UPD. was wrong on smaller, it's actually much larger

lostmsu··on Covid can wake up a slew of dormant viruses inside you
Any weakness can
lostmsu··on A year of fighting scrapers on my 1.5 million-page website
Huh? I regularly ask codex to not just summarize or extract a crux of published papers, but sometimes to do a topic directed search and give a comparison table. If you aren't doing that, do you see the problem?
lostmsu··on The West's Demographic Math No Longer Adds Up
> She says she worries that becoming a mother would mean giving up the life she has now.

The lady even told you the reason, but you still brought up the climate change, religion, and social media.

lostmsu··on Can Intel finally beat ARM on performance per Watt?
When you test matmul these days you want bf16 max. FP64 is a niche workload 100% unused by 95%+ users.
lostmsu··on 2027 memory capacity is reportedly sold out
Sure, start sabotage with this one then https://oceancos.com/ocean-cold

Salmon is not too important anyway.

lostmsu··on Monitors for Work
First, it wasn't an ultrawide.

Second, I used my own custom window manager on Windows: https://github.com/StackWM/

lostmsu··on A year of fighting scrapers on my 1.5 million-page website
Because the effort you are talking about is completely unnecessary. Imagine every grocery store would require you do a little dance when you buy a carton of milk. Your vision needs some narrowing (which I believe will justify the parent point), because as-is it has obvious counterexamples.

I believe your problem is that "effort" is unspecified. "some effort" would make the statement correct, but some effort does not justify arbitrary effort, therefore you have no point here.

lostmsu··on A year of fighting scrapers on my 1.5 million-page website
The situation is getting worse in some OSS communities too. I go to a bugtracker just to read the discussion on the issue I am facing, often to understand what's the roadmap to fix it if any, and some of them immediately demand me to enable JS and solve a captcha. Even GitHub doesn't do it!
lostmsu··on A year of fighting scrapers on my 1.5 million-page website
> Do you think the LLM reads every page on the internet before generating your answer? Of course not.

The way LLMs are trained the answer is of course yes, but they don't remember them all exactly, of course.

lostmsu··on A year of fighting scrapers on my 1.5 million-page website
> at worst devastating

This is such a low bar literally anything would clear it. Even tomato farming.

lostmsu··on A year of fighting scrapers on my 1.5 million-page website
From Google's perspective you might as well be a bot because when you search there you don't see ads. If you don't have adblock installed, your parent comment applies because you get worse experience as LLMs don't serve ads.
lostmsu··on 2027 memory capacity is reportedly sold out
Every produce storage facility going to have them, won't it?
lostmsu··on Why Normal People Aren't Using AI Agents
> They just use it without knowing

No they don't, not efficiently (if you are referring to indirect access).

"Normal" is synonym to average in this case.

lostmsu··on Our world is full of slop
I feel like this article severely underestimates amount of slop in the world. Almost everything on the Internet is slop.

The only real non-slop is manufactured/built/grown stuff and a relatively short list of art works.

lostmsu··on Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
They were trained to do that specifically, don't you think?
lostmsu··on Trump's tech ties blamed for AI inaction by MAGA Republicans and Democrats
What exactly do you want to regulate?
lostmsu··on Nvidia’s Vera Whitepaper Has a Thread Loose
You forgot their tensor core performance numbers "with sparsity".
lostmsu··on Democratic lawmakers propose 3-year moratorium on new data centers in Oregon
There are plenty other states
lostmsu··on Unified Neuro Symbolic Engine AGPL 3.0
Get human help
lostmsu··on COVID can wake up a slew of dormant viruses inside you
Any sickness does that
lostmsu··on Muse Spark 1.2 (Xhigh) Intelligence, Performance and Price Analysis
Not at the frontier. Good luck next time!
lostmsu··on Zed DeltaDB
I don't know. I've been giving Zed a chance for almost half a year now and it just doesn't seem to be going the way I would want an editor to go.
lostmsu··on Monitors for Work
In 2016 I bought a $700 40in curved 4k monitor for my job at Amazon. It's funny they probably spent over $1500 on a laptop I almost never used but did not want to pay for a decent monitor.
← PreviousPage 4 of 34Next →