Apple's rumored M7 Ultra targets 1.5TB and Blackwell-class AI performance
tomshardware.com
tomshardware.com
Similar story with Qualcomm but to a lesser degree (both on the CPU and GPU side). I wonder why that is.
Lack of priority? Legitimately harder problem to solve? Experience from mobile scaling to desktop differently than CPU experience?
(My take. happy to be corrected by someone with chip design experience who can comment.)
A high IPC is much easier to achieve when the instructions have a fixed-length encoding, so this RISC principle followed from their main choice.
Another user though pointed out that they didn't actually design the GPU until the A11, so perhaps it really is just a lack of in house experience.
Doesn't the argument work fine on the CPU side? Apple doesn't seem to be hurting for lack of a Threadripper competitor.
https://arxiv.org/html/2502.05317v1
Apple vs. Oranges: Evaluating the Apple Silicon M-Series SoCs for HPC Performance and Efficiency
"Apple's M-series GPUs offer massive performance-per-watt, scaling efficiently from roughly 5W in base chips to around 40-50W in top-tier Max chips. This efficiency generally sits between 200 and 250 GFLOPS per Watt."
Sure. nVidia GPUs have higher performance. But they do it with 450W+ a.k.a. 10x the power
Nvidia's chips lose very little performance when they're severely power limited. In fact, that's why Nvidia's Professional versions of their GPUs are typically running at under two-thirds of their consumer equivalent's power draw.
Nvidia's cards are primarily designed for their professional applications where the power consumption is lower, and then they just juice up the consumer cards deep into the territory of diminishing returns just so they win some benchmarks.
On a performance-per-watt basis, Apple is still behind an Nvidia card with the juiced up power consumption, and when you dial back the power draw, Nvidia cards are miles ahead of Apple's silicon for performance-per-watt.
Do you have a citation on Apple vs nVidia performance-per-watt? I'm not aware of any benchmark that shows nVidia with a better performance-per-watt. Higher performance, sure, but performance-per-watt, not that I'm aware of.
Right now the M5 Max is one of the best laptop compute GPUs for blender rendering and doesn’t do too badly on raster tests either. But it operates at a fraction of the equivalent NVIDIA parts.
It’s more that Apple doesn’t provide a competing tier of device as the high end NVIDIA GPUs. But within the tiers they do provide they’re fairly competitive.
I’m not sure if I can measure IPC and cache misse latencies on my machines, but I’m sure latencies are pushing IPC down.
Whether or not this thing is commercially available though, they basically have no choice but to at least keep up the R&D spend, or they'll be even further behind once RAM availability eventually eases up.
Blackwell came out in 2024 on n4p. This article is claiming that Apple hopes to get into the same ballpark of performance as a Blackwell GPU with an M7 Ultra, which at the absolute earliest would release in 2027, but more likely 2028 or 2029, and would consist of two absolutely massive M7 Max GPUs stapled together, for a total die area bigger than a 5090, but on a smaller more advanced node.
Honestly, if they can't release something that big by 2028 that's at least competitive with a 5080, it'd be extremely embarrassing for them.
For AI purposes that may change but I doubt it.