HNHacker News
TopNewBestAskShowJobs

pama

3,780 karma · joined October 11, 2010

submissionscomments
pama··on M5 Ultra Mac Studio Review
But what about builds that combine 8 of the 5090 with infiniband between boxes? Wouldn't that be comparable to the mac in terms of price and potentially beat it by a lot in terms of performance for the large MoE? I understand the space/heat/noise considerations, but price wise it may still not make as much sense as people think. (Agreed that it is hard to get the NVIDIA hardware and the 6000 pro are priced less competitively).
pama··on Google's Open Agentic Orchestrator
Not GP, but you start with Kubernetes…

> You need a Kubernetes cluster, ko (brew install ko), a container registry your cluster can pull from, and a reachable Agent Substrate Control API (in-cluster default: api.ate-system.svc.cluster.local:443).

> make deploy AX_IMAGE_REPO=<your-registry>

> This deploys Redis, then builds and deploys the control plane images with ko. Everything lands in the ax-system namespace.

pama··on How OpenAI Used Its Own LLMs to Design Its Jalapeño Chip
Having worked with people doing bringup of specialized chips, I am awed at how the world has changed.

> When the first chips came back from the foundry in May, the team pointed its internal AI models at designing software to run benchmarks such as SemiAnalysis’s InferenceX. On DeepSeek’s multi-head latent attention kernel benchmark, performance climbed from 0.31 percent of the theoretical ceiling (set by the chip’s compute and memory bandwidth) to 88.94 percent in roughly 40 hours. Ho says this result is repeatable, so the time between when foundries deliver the first chips and when production ramps up can be reduced. “All our schedule assumptions are going to be based on the fact we have this capability now,” he says.

pama··on How to Write with an LLM
How I write with an LLM: for each page “suggest up to 10 word changes to increase clarity.”
pama··on Introducing System One Models and Jev
Is there a downloadable technical report somewhere?
pama··on Navier-Stokes Announcement
This person knew they did not prove the Collatz conjecture and others independently figured it out within hours. Not sure this is at all relevant, other than pointing out how trivial it is for the community to understand errors in lean4.
pama··on On the Navier–Stokes Millennium Prize Problem
I agree with search, which is what mathematicians also use over longer periods of time. But it is not brute force search (and neither is alphago’s search or modern stockfish, though both still search at depth and speed higher than typical human).
pama··on Navier-Stokes Announcement
Not sure what you mean. Here is what happened in that case: https://news.ycombinator.com/item?id=49137060#49140177
pama··on Litelm: LiteLLM Without the Bloat
You misunderstood. This new project has 2,900 LOC. Maybe the spelling change is too subtle.
pama··on Navier-Stokes Announcement
Perhaps you did not understand the Fermat theorem proof announcement/repo or the link. The 13 million lines did not use any external, possibly not honest libraries, as the proof eventually only used the fundamental axioms. So for the Fermat theorem formalization, no open open questions remain.
pama··on Astra for Coding: Why Are We Doing This Again?
I had good luck with Kevin Lin’s tip for Astra: “Can you radically simplify the implementation?”
pama··on On the Navier–Stokes Millennium Prize Problem
You jest and that is OK. Brute force search is not something you can do over math problems of that difficulty or anything with combinatorial complexity.

To me it feels closer to taking the top 10k human mathematicians on a large retreat for a year and having them self organize to collectively solve this problem—not kids and easter eggs.

pama··on On the Navier–Stokes Millennium Prize Problem
As far as I understand the 10k agents worked on the proof. The lean formalization came later and was easier/faster than getting the proof.
pama··on On the Navier–Stokes Millennium Prize Problem
If you work with distributed systems, you still call that scenario a success. On the other hand, if the 3/4 of agents going off the rails bring down the whole mission, that is a failure. The latter would have been my guess with current models scaling to 10k agents.

I am not an expert in lean4, but I could follow parts of the high level lean definitions of the problem statement in the repo. A lean bug would be a fun scenario; I am certain this proof will receive the deserved scrutiny, and if it uncovers a bug, it will make the story even more exciting. It is extremely unlikely to be the case, however, because the 10k agents working on the proof didnt use lean, so it would have to be a math logic error that translates to a lean bug—perhaps something the agents picked up during training?

pama··on I'm sorry, you're not going to die from an AI-engineered supervirus
I am not sure what you mean here. There exist plenty high-concern biological threats that dont need any AI help. Human oncovirus design is low on my concern list (immunity is diverse), and in any case it does not need AI—rather labspace. Disgruntled high schoolers or undergrad chemists can do way more damage from readily available materials without AI and without research delay. As can nature (or amateur biologists without AI) by mixing bats with their animal of choice and waiting a while. If/when any scary global events like covid happen again, I sure hope we have true superintelligence to help us navigate it quickly.
pama··on How An AI math breakthrough ignited a controversy
Other than the undeniable breakthrough in math, the important point is the ability to orchestrate 10k agents to productively work on a single problem, which creates options:

> OpenAI, meanwhile, says its experience with Navier-Stokes could open the door to solving puzzles with more practical relevance. “We are now able to spend millions of dollars on a problem that we really care about and that really matters: developing new materials, finding cures to diseases,” Bubeck said. “All of those things that we have been talking about for a long time—now they seem to be at our fingertips.”

pama··on On the Navier–Stokes Millennium Prize Problem
They managed to solve a problem that was beyond current human ability.
pama··on On the Navier–Stokes Millennium Prize Problem
Not only that, but it used 10k agents coherently over 88 hours to come up with the proof. This is a significant advance.
pama··on uv: Deduplicate all files in the wheel cache
So at 3 million different files you have a 98.3% chance of a hash collision. Wouldnt that cause problems in real datasets?
pama··on C++26: Standard Library Hardening Experiments
30 years late, but I will take it. Contracts look useful and less messy than exceptions.
pama··on Show HN: The load-bearing vocabulary of Claude
Oh, the use of “load-bearing” is a pretty clean bug in the system context of Claude code. It can be fixed trivially. The overuse of particular language constructs, including the one you suggested, is a more interesting problem; it is fixable, but might involve considerably more effort.
pama··on Show HN: The load-bearing vocabulary of Claude
Interesting use of Claude’s construct! Your phrasing might also suggest that it resolves something else, which is left ambiguous, whereas your other construct would be less likely to suggest that idea (unless you explicitly appended a contrastive clause for emphasis).

The bots are subtle. From your example in your GP comment (and in part depending on the surrounding context) I would expect that an empty list would be less likely the subject of: “the list contains no string”, than in “the list doesn’t contain strings”.

I might be over interpreting the intentionality of Claude in using this construct, however, these models were pretrained by reading so much more than any human, they learn to handle language differently than most humans.

pama··on Show HN: The load-bearing vocabulary of Claude
> "It <verb>s no <noun>" instead of "It doesn't <verb> <noun>"

The meaning is often different in these constructs. Consider: “Claude answers no questions” vs “Claude doesn’t answer questions”. The first could be a bot or a politician avoiding the substance, the second could be a broken UI or a politician cancelling the QA of a press conference.

pama··on Debian polls its developers on AI: permit or ban?
> Is the energy usage so different between local and cloud inference?

In throughput mode for agentic loads, the energy usage (tok/s/MW) of the new NVidia Vera Rubin hardware is 30x lower than that of the B300 and perhaps 450x lower then the H200 was, which in turn is hundreds of times lower than the inference for single users at home in any non-data-center hardware. It feels like comparing the momentum of an ant to the momentum of an elephant.

pama··on New Mac mini, featuring M6 and M5 Pro
> I'm kind of lost on what to do with it.

Personal assistant (e.g. openclaw) in the mac ecosystem. You need external models to power it ofc, but cool to have all the Apple integration and local network in your assistant. The claw can be addressed in multiple ways and can command your local hardware if you so wish.

pama··on Why your local LLM feels dumber than it is
Without doubt, dsv4-flash-0731. Original weights; needs two connected DGX.
pama··on Emacs 31.1 will release on 8/24
Thanks! I hadn't tried it. I've started using it today; I limited a couple options about dired-local and images in kitty due to concerns from upstream security audits, but otherwise looks fantastic.
pama··on Emacs 31.1 will release on 8/24
Thanks! Agreed ghostel looks fantastic. I still wish there was an in between a full TUI and a full non-interactive version, but switching to ghostel means I can fully live within emacs again.
pama··on Emacs 31.1 will release on 8/24
Similar story here (user since early 90s). I wish more AI harnesses could also have a dumb-terminal mode in addition to the fully non-interactove mode, so they could all stay naturally searchable/copyable in simple comint buffers. Even the best current terminal emulators within Emacs are not great to recursively handle Emacs (or modern AI TUI) well enough, so I occasionally also use a shell putside Emacs.
pama··on AliExpress runs silent WebAudio fingerprinting that breaks Bluetooth multipoint
Another reason why Lockdown mode on iOS is your friend.
Page 1 of 32Next →