Microsoft acquires twice as many Nvidia AI chips as tech rivals
ft.com
ft.com
"Omdia analyses companies’ publicly disclosed capital spending, server shipments and supply chain intelligence to calculate its estimates."
That could explain a large part of the gap.
Edit: it looks like Microsoft announced their own last year, but I can imagine they may be behind the curve in capability and scale out compared to the others
Intel is better at chip design than any of those companies. They spent a lot of effort coming up with a very clever chip that competed well against the current generation of Nvidia chips, while still running your old x86 codes.
Nvidia continued increasing memory bandwidth, and nobody cared about Knights Whatever.
What will they accomplish with the things? Why even think about that part? Probably AI. Selling premium GEMMs, what a trick. Bah. Hopefully TSMC got a really good cut, they are at least doing some interesting engineering.
An estimated count is newsworthy for the journalists and the readers because it's an indirect proxy for outsiders -- who are not privy to the internal plans of FAANG companies -- to try and figure out what's happening. Basically trying to "read the tea leaves" of the AI industry.
Demand exceeds supply. NVIDIA has limited number chips to sell and TSMC factory time is overbooked. In the current zero-sum situation, NVIDIA picking and choosing who to sell to may be a signal of something. And/or Microsoft/OpenAI's willingness to spend billions on 2x the NVIDIA chips is a signal of something.
https://fortune.com/2024/02/21/nvidia-earnings-ceo-jensen-hu...
https://fortune.com/2024/09/12/nvidia-jensen-huang-ai-traini...
Fascinating, I had no idea this was even a thing. Simultaneously badass team green is far ahead and also a bummer to be artificially limited / segmented.
I wonder if the upcoming 5090 core will mostly be a fuse-intact 4090. I imagine nearly all of NVs current focus is on H200 and Blackwell and whatever else is in the pipeline rather than these "silly" little gamer cards which bring in comparably trivial financial resources.
/me *cries a tear*
Maybe, but it's also possible that it wouldn't do anything. Other Nvidia boards (like the Tegra in the Switch) also come with arbitrarily disabled "dark silicon", but enabling the extra SOC hardware only causes the board to crash when using everything at once. It wouldn't surprise me if this was a binning measure, even though I also wouldn't be surprised if it was an arbitrary limit.
https://x.com/cognitivecompai/status/1868399108924592391
https://x.com/cognitivecompai/status/1868401738706993301
From my research it seems that restoring full GPU capabilities by repairing or circumventing a deliberately blown eFuse on an NVIDIA AD102 die is, for all practical purposes, impossible.
While Nvidia might have won the training market, it’s inference where the real money is.
Additionally - TPUs are completely useless if AI goes out of style, unlike CUDA GPUs. The great thing about Nvidia's hardware right now is that you can truly use the GPU for whatever you want. Maybe AI falls through in 2026, and now those GPUs can be used for protein folding or crypto mining. Maybe crypto mining and protein folding falls through - you can still use most of those GPUs for raster renders and gaming too! TPUs are just TPUs - if AI demand goes away, your dedicated tensor hardware is dead weight.
https://finance.yahoo.com/news/microsoft-stock-receives-rare...