HNHacker News
TopNewBestAskShowJobs

_m5co

1 karma · joined August 19, 2026

submissionscomments
_m5co··on Nvidia H100 GPUs: Supply and Demand
Thank you
_m5co··on Nvidia H100 GPUs: Supply and Demand
I appreciate it, thanks for the comment
_m5co··on Nvidia H100 GPUs: Supply and Demand
I agree. (I'm the author) Touched on that briefly here https://news.ycombinator.com/item?id=36955403. Need help with that research; please email - email is in profile. Had a section on it in early drafts; didn't feel confident enough; removed it.

Would be good to have more on enterprise companies like Pepsi, BMW, Bentley, Lowes, as well as other HPC uses like oil and gas, others in manufacturing, others in automotive, weather forecasting.

_m5co··on Nvidia H100 GPUs: Supply and Demand
Thanks! (Author here, see other work in my HN submissions and comments)
_m5co··on Nvidia H100 GPUs: Supply and Demand
(Author here) I'd be interested in writing about this in the future. I need help though because I don't know people in those spaces. Email is in my profile. I had a section on this in early drafts but removed it as I didn't feel confident enough in my research.
_m5co··on Nvidia H100 GPUs: Supply and Demand
(I'm the author of the linked post)

Yes, much needed.

Here's a list of possible "monopoly breakers" I'm going to write about in another post - some of these are things people are using today, some are available but don't have much user adoption, some are technically available but very hard to purchase or rent/use, and some aren't yet available:

* Software: OpenAI's Triton (you might've noticed it mentioned in some of "TheBloke" model releases and as an option in the oobabooga text-generation-webui), Modular's Mojo (on top of MLIR), OctoML (from the creators of TVM), geohot's tiny corp, CUDA porting efforts, PyTorch as a way of reducing reliance on CUDA

* Hardware: TPUs, Amazon Inferentia, Cloud companies working on chips (Microsoft Project Athena, AWS Tranium, TPU v5), chip startups (Cerebras, Tenstorrent), AMD's MI300A and MI300X, Tesla Dojo and D1, Meta's MTIA, Habana Gaudi, LLM ASICs, [+ Moore Threads]

The A/H100 with infiniband are still the most common request for startups doing LLM training though.

The current angle I'm thinking about for the post would be to actually use them all. Take Llama 2, and see which software and hardware approaches we can get inference working on (would leave training to a follow-up post), write about how much of a hassle it is (to get access/to purchase/to rent, and to get running), and what the inference speed is like. That might be too ambitious though, I could see it taking a while. If any freelancers want to help me research and write this, email is in my profile. No points for companies that talk a big game but don't have a product that can actually be purchased/used, I think - they'd be relegated to a "things to watch for in future" section.

_m5co··on The GPU Song [video]
I helped make this! It’s very very niche, but if this happens to be your niche then I think you’ll enjoy it quite a bit.