Most interested to see their inference stack, hope that’s one of the 5. I think most people are running R1 on a single H200 node but Deepseek had much lower RAM per GPU for their inference and so had some cluster based MoE deployment.
Most interested to see their inference stack, hope that’s one of the 5. I think most people are running R1 on a single H200 node but Deepseek had much lower RAM per GPU for their inference and so had some cluster based MoE deployment.
But yeah, would be interesting to see how they optimized for that.
H20 is the next iteration of the "sanctions" that I believe also limited the "cores" but left the on-chip memory intact, or slightly higher (from the new generation).
> We believe DeepSeek has access to around 10,000 of these H800s and about 10,000 H100s. Furthermore they have orders for many more H20’s, with Nvidia having produced over 1 million of the China specific GPU in the last 9 months.
that's as dumb as saying coca cola have acccess to all offices of Berkshire Hathaway.
likewise, all comments praising deepseek history are also misleading as the company barely exists for a year.
everything is opaque marketing being repeated. just drop the off topic bla bla bla and focus on the facts and code in front of you.
thanks for coming to my ted talk.
If you wouldn't mind reviewing https://news.ycombinator.com/newsguidelines.html and taking the intended spirit of the site more to heart, we'd be grateful. The basic idea is to make your substantive points thoughtfully, regardless of how wrong anyone else is or you feel they are.
Claiming they have access to 5x this amount is not such a bold claim?
What claims from the semianalysis article do you think are false? And based on what evidence?