Azure announces new AI optimized VM series featuring AMD's flagship MI300X GPU
techcommunity.microsoft.com
techcommunity.microsoft.com
Microsoft has the know-how from hardware and software, drivers, APIs, firmware the resources, it's a major AMD customer in Cloud and consumer devices.
Do people think Microsoft and AMD will watch Nvidia cannibalize the market snd Microsoft writing cheques for whatever Nvidia demands?
It's like people forget a major economic rule: when margins are high you will attract competition.
If pressed, sure, people will admit that these may not be entirely imaginary entities, but if listing technologies or platforms then “oops” they’ll just forget.
The best example I saw was a “poster” of cloud big data technologies. Started with Amazon S3, went through Google Bigtable, and then in the corners had companies so small that their own marketing page is the only search result. No mention of Azure anywhere.
There were dozens of logos on that page representing companies with annual revenues smaller than what one my customers spent on a single Azure Storage Account by accident.
https://arstechnica.com/information-technology/2023/11/bing-...
Meh, the world of developers as a profession, is much bigger than some angry activists on Twitter and HN smelling their own farts in their echo-chambers and venting on their blog.
HN also likes to forget that enterprise also exists. And that IBM or SAP exist. Or that most of the tech that makes the world go round is very outdated and not the latest hip cool stacks from the start-up world.
Sometimes it's just another echo chamber where people prioritize emotions and feelings over facts.
However, I don't think HN forgets they exist, I think a lot of HN users are those enterprise people who want to use better tools.
That's like >90% of SW developers in the world.
You can do everything with one vendor, and some customers really love that. Ours certainly do. They don't even bother trying to audit us along those lines anymore. We make it as simple as possible at every level. Offering a successful B2B SaaS experience is more about the biz interactions and less about the nerd crap. If you found a way to make your internal tech more important than the biz, you likely have a problem.
Every time we talk about databases, I do a CTRL+F and check for 2 things: Is SQLite covered and are we being assholes about SQL Server for no good reason. SQLite gets a good rap on HN, so I more often prioritize awareness around the red-headed stepchild of RDBMS. Their hyperscale offering is end-game for 99.99% of business use cases.
I always end up at "How much money, time and frustration are your principles worth?"
This is totally correct if you are doing small scale stuff in the cloud or just random "enterprise" for sub-thousand person companies. Azure is playing a distant third in reliability, maybe even fourth behind Alibaba.
HN probably leans towards people who build things with the cloud, whereas it sounds like your customers are people who want to pay people to run things in the cloud for them.
The talking heads on CNBC would probably mention their synergy of "software, drivers, APIs, firmware the resources" but as a user of their end results - once bitten, twice shy.
There have already been several ChatGPT outages caused by a lack of compute capacity. Azure literally cannot buy Nvidia GPUs fast enough to satisfy customer demand. They are forced to buy alternatives here. There's no real decision to make here.
https://news.microsoft.com/2023/09/14/microsoft-and-oracle-e...
Specifically for the parts of Bing that are rebranded OpenAI https://www.oracle.com/news/announcement/oracle-cloud-infras...
https://www.oracle.com/news/announcement/oracle-cloud-infras...
Pure PyTorch mostly works OK, but some libraries implementing crazy optimized, hand written kernels and such will have some trouble.
So (for instance) maybe you can run an LLM in a particular PyTorch framework, but flash attention 2 doesn't support your AMD card, so performance and memory use takes a hit.
Or maybe the library works on an Intel XPU with like 5 changed lines in the entire library (rename "cuda" to "xpu"), but no one bothered to add it, or maybe the dev doesn't even want to support the PR.
If I have an NVIDIA GPU, the only PyTorch backend I can promise even "somewhat works on my machine" is CUDA.
They benefit from competition, not from bending to one vendor.
It is all sold out, for years. You can't sell something you don't have.
https://fortune.com/europe/2023/10/04/amd-lisa-su-nvidia-roc...
So yea, in my eyes, AMD beat Intel. I don't think I need any more evidence than that. ¯\_(ツ)_/¯
(hello!)
There are a lot of moving parts around all of this. AMD was still dealing with their fab breakup and fallout.
They basically cut all the fat and some muscle to keep pushing forward so I agree with the other poster that now is the time to focus on software and growth.
https://twitter.com/sama/status/1724626002595471740
ROCm has also made a lot of advances in recent times.
https://www.databricks.com/blog/training-llms-scale-amd-mi25...
It was designed as a combined CPU/GPU for supercomputers, with shared memory. But then the AI craze hit, so AMD spun a variant into a pure GPU AI accelerator real quick, which they could actually pull off because the GPU silicon is modular.
...So thats why it cost a fortune. Its really a jury rigged HPC product.
Intel is taking a slightly different approach, and is going for "PyTorch compatible."
You will hear endless negative anecdotes about ROCm/OpenVINO, but they both do seem to be getting better with each update.