To those commenting about "no moat" remember CUDA is a huge part of it, it's actually HW+SW and both took a decade to mature, together
To those commenting about "no moat" remember CUDA is a huge part of it, it's actually HW+SW and both took a decade to mature, together
Gaudi2 was actually announced 2 years ago and is 7nm like the A100 80Gb it was meant to be competitive with, Gaudi3 later this year is probably going to be the inflection point as that ramps
The cost is like 1/3
https://www.intel.com/content/www/us/en/newsroom/news/vision...
- Intel acquired Habana in 2019
- Habana launched Gaudi2 in 2022
- only in H2 2023 Habana enabled FP8 which delivered around 100% improvement in time-to-train
On the rest I believe you but markets don't move based on single individual's/company's data points
We are about to drop stable diffusion 3 which is the best image model out there (https://x.com/EMostaque/status/1764941367682256950?s=20) with similar architecture to Sora by OpenAI that can be used for any modality.
We have hundreds of millions of downloads of our models so are looking for big scale as we move to every pixel being generated & this stuff goes from research to mass deployment.
2024: Nvidia's B100 TSMC 3nm (?)
2024: Intel Gaudi3 TSMC 5nm (*)
2023: AMD MI300X TSMC 5nm/6nm
2022: Nvidia H100 TSMC 4N
2020 Nvidia A100 TSMC 7nm
(*): performance critical chiplets at least.