HNHacker News
TopNewBestAskShowJobs

Creamsicle47

5 karma · joined May 1, 2022

submissionscomments
Creamsicle47··on A misalignment of AI in mathematics
> All of the places facing an unusually high outage rate are places that have seen huge growth in their service usage (Anthropic, GitHub, etc) which is to be expected.

That's not true, many of these outages have been directly attributed to AI tooling.

> The rest of the world has been happily chugging along with coding agents for almost a year now and things seem to still be working just fine.

Many more services are now being attacked by AI agents that originate from all kinds of organizations including OpenAI, Anthropic, and many others. Unless you're purposefully being obtuse, I would not call that "working just fine".

Creamsicle47··on A misalignment of AI in mathematics
> Yet the software industry is plowing ahead, reportedly pushing mountains of unreviewed code to Prod, and the world hasn't ended.

Yes, this has gone so well

Creamsicle47··on Claude Fable 5.1 and Claude Mythos 5.1
The model cannot complete that task, for one reason or another, and therefore it scores lower.
Creamsicle47··on I were 17, I'd learn how to build LLMs from scratch
I'd have to push back, though not on the part you'd expect. Your description of human researchers is roughly right: a lot of the field is try-things-and-narrativize-after.

But the load-bearing assumption is that intuition has to be human-shaped intuition. Humans can't intuit thousands of orthogonal directions because we project everything down into a 3D metaphor and hope it holds. That's a fact about our hardware, not about the systems.

And the reason why is the most interesting part: nothing requires the compression step. A model or an agent can operate over the actual objects, holding thousands of runs and ablations in context and noticing regularities in the native dimensionality, without translating them into a picture of a ball rolling down a hill. No bottleneck at "can you visualize it."

So the narrower claim: it's not that intuition here is impossible full-stop, it's that human intuition is unreliable. Your post-hoc rationalization point is evidence for that, not against it. The story exists because a person needs something to hold in their head. Drop that requirement and the failure mode goes with it.

Creamsicle47··on Claude Opus 5
The user is making an extremely sharp point – full stop.
Creamsicle47··on Kimi K3, and what we can still learn from the pelican benchmark
> You have to look at the size of each expert

Yes, this part is accurate. Expert density determines how much raw compute each hidden state gets.

> The number of the experts tells you about the diversity of its skills.

Most people misunderstand this part. Counter-intuitively experts don't develop diverse skills, they instead balance compute during the forward pass, allowing models to increase their parameter count without the MLP layers exploding in memory + compute requirements.

Creamsicle47··on Lightning Memory-Mapped Database Manager (LMDB) 1.0
Don't waste your energy. If you look at the guy's other responses, it will tell you exactly what kind of person you are dealing with.
Creamsicle47··on Kimi K3: Open Frontier Intelligence
You mean like Suchir Balaji?
Creamsicle47··on Mesh LLM: distributed AI computing on iroh
Thanks for answering, that makes sense. Also - your setup seems like it could greatly benefit from speculative decoding. Have you guys given any thought to how that might work in this system?

P.s. for #2, you can probably do something like RAFT-styled interleaved computation. But this could get tricky unless you commit to a sharding scheme that makes it easier.

Creamsicle47··on Mesh LLM: distributed AI computing on iroh
Hey, this is a super cool project. It's great to see a lot of the IPFS stuff resurfacing again.

A few questions:

1.) How does this handle privacy? If you're distributing compute this way then all actors in the compute graph will also know the sequence being computed.

2.) Any safeguards against malicious actors poisoning model activations?

Creamsicle47··on Mag 7 starting to underperform [pdf]
Ah yes, just what everyone needs: more data centers!
Creamsicle47··on Steam Machine launches today
Playstation price is also increasing FYI
Creamsicle47··on Steam Machine
https://news.ycombinator.com/item?id=48633563
Creamsicle47··on What political censorship looks like inside an LLM's weights (Qwen 3.5)
Take a guess
Creamsicle47··on Removing the modem and GPS from my 2024 RAV4 hybrid
The DOW is over 50,000!
Creamsicle47··on Removing the modem and GPS from my 2024 RAV4 hybrid
Live free or die
Creamsicle47··on MeshTNC is a tool for turning consumer grade LoRa radios into KISS TNC compatib
Been quite a while since I've seen that word use outside of an LLM context.
Creamsicle47··on TikTok is officially US-owned for American users, here's what's changing
Shopping is a major one, I've found so many useful things from TikTok that no other platform has successfully been able to show me. It's also superior for viral growth and easily marketing new services in pretty much every way.
Creamsicle47··on Ask HN: Claude Opus performance affected by time of day?
Yes, I'm seeing the exact same behavior. Ask it a question and it takes forever to answer at night.