AMD 'Strix Halo' Ryzen AI Max+ Debuts with RDNA 3.5 Graphics and Zen 5 CPU Cores
tomshardware.com
tomshardware.com
If it runs tinygrad at speed(a lower bar developmentally) I might get one.
Is there a model benchmarking site where you can select varying degrees of models by source code and see how they perform on different hardware. It would assist people to evaluate whether or not a specific piece of hardware is good for the jobs that they want it to do.
Not that I'm aware of (at least based on real benchmarks), but it's something I've been noodling about building, together with with some other associated data that can be helpful when wanting to select a model. Glad to hear I'm not the only one wanting it :)
https://www.digitaltrends.com/wp-content/uploads/2025/01/3d-...
Why can't companies just include absolute numbers in their comparisons...
Non-thumbnail version of the chart: https://www.digitaltrends.com/wp-content/uploads/2025/01/3d-...
100% chance its chewing through at least 50% more power to achieve the result.
Infact based on their TDP guidance, it goes up to 120w, which is more than double M4. But we don't know what the configuration was for this benchmark. We also don't have great numbers for M4's power consumption either.
Then you throw in the fact 120w TDP from AMD is not actually a power consumption figure... and it's all made up.
https://www.reddit.com/r/macbookpro/comments/1hj3m0p/m4_max_...
https://forums.macrumors.com/threads/m4-max-eats-battery.244...
M4 is likely more power efficient, but not 2x.
It would be behind the M4 Max. It’s also over double the wattage of the M4 Pro to achieve these numbers.
256 bits * DDR5-8533 is a pretty big step up from any other x86-64 laptop or SFF and should be a pretty huge help for anything graphics or bandwidth intensive, like LLMs.
Gosh that name is a mouthful
https://en.wikipedia.org/wiki/Strix_(mythology)#Greek_origin...
"AI" seems to have replaced the segment number.
The + is because it's the top-end model of the lineup.
Not sure what's Max or Pro about it though.
Max is the segment number i.e. it's "Ryzen 11" (the other 300 series SKUs they announced are Ryzen AI [5|7] 3xx). Weirdly though there are no Ryzen 9s so maybe it's really just a rebrand of 9.
The Pro just means it has management and security features for enterprise customers.
BTW, here's the meaning of "Strix": https://en.wikipedia.org/wiki/Strix_(mythology)
Secure processor, shadow stacks, secure boot, hardware asset trackability. Enterprise stuff.
That was the good part, lol.
Bandwidth is more or less on par with the M4 Pro, and it supports up to 128GB.
Furthermore, during the chip shortages of the last few years, AMD was actually selling broken PS5 silicon for use as a normal Windows PC[0]. If there were restrictions on selling APUs above a certain performance level, then this PC wouldn't exist.
The OEMs buying APUs to use in laptops and SFF desktops were more interested in cutting costs than boosting graphics performance. Users who want better 3D performance can buy a higher end laptop with a discrete GPU and juicier profit margin.
Doubling the memory width (and tripling the bandwidth) helped Apple's GPU performance substantially and should do the same for AMD. Which means that a larger fraction of the laptop market should consider it "good enough" and still have a reasonable TDP to avoid the 2" think laptop that last for less than an hour on battery while sounding like a hair dryer.
Zen 5 CPU, RDNA 3.5 GPU, and XDNA 2 NPU. No word on process nodes.
8 GHz x 32 bytes = 256 GB/s
This has been known for a long time.
What annoys me is that AMD does not say whether the Zen 5 cores of Strix Halo have full vector processing pipelines, like Granite Ridge and Fire Range, or they have the narrow pipelines of Strix Point and Krackan Point.
Maybe even a SFF sized motherboard that allows CUDIMMs, which is a nice fit since each CUDIMM is 128 bits wide.
How far in the future? I don't need another laptop, but would be nice to have a box to run local llms on. If these things can run LLMs at a decent clip then this would be sort of a "shut up and take my money" situation.
EDIT: Oh, I see they're calling the laptop a workstation.
[1] https://liliputing.com/hp-zbook-ultra-14-g1a-mobile-workstat...
https://www.pcworld.com/article/2567865/hp-z2-mini-g1a-packs...
I've heard similar rumors of similar SFFs from framework, system76, and similar VARs.
The Strix Halo does seem pretty compelling for those that want less volume, power, and money than a discrete GPU. I'd love a small SFF with either 128GB ram (some parts will have the ram in the package)) or two CUDIMM slots.
Generally SFFs use laptop parts, but tend to be out a few months later than the equivalent laptops.
I expect similar with the Strix Halo/AI Max.
HP announced a HP Z2 mini g1a, which is bigger than a NUC, but I believe still considered a SFF:
https://www.pcworld.com/article/2567865/hp-z2-mini-g1a-packs...
However one big surprise was that the Halo 395 chip runs Llama 3.1 70B-Q4 2.2x times faster than a RTX 4090 24GB. Anyone have any details? The slide mentions seeing AMD endnote SHO-14 for details.
Maybe 70B-Q4 doesn't fit in 24GB?
Not so impressive 8-(.
The bandwidth should mostly help the GPU performance for games or LLMs, but not random desktop apps.
100fps is 10ms for a frame (we now know 60pfs is not enough, fps must never drop below 75-80fps, but I would really target 100+ fps)
Laptops with 120W cpus aren't thin or light.
Edit: per article these are positioned from 45w up
Right, the CPU.
"All of the AI Max chips have a 55W base TDP, but also a configurable TDP that ranges from 45 to 120W to unleash more horsepower in designs that can handle the thermal output."
I wonder if at this point, NVMe further becomes a bottlenecking factor?
Doubly so that for the last decade or so memory sizes haven't increases. Seems like the vast majority of machines these days are 8-16GB ram and have a max of 64GB ram unless you are buying workstation parts.
How many times a second do you need to load 100% of ram from storage?
At least the strix halo supports 128GB of ram.