The M2 Ultra has a 800 Gbps bandwidth, still greater than M4 Max. For running LLMs, this makes a huge difference.
P.S: Nvidia's 4090 has a 1000 Gbps bandwidth.
The M2 Ultra has a 800 Gbps bandwidth, still greater than M4 Max. For running LLMs, this makes a huge difference.
P.S: Nvidia's 4090 has a 1000 Gbps bandwidth.
Also, as a customer, I have become somewhat annoyed by Apple's naming scheme on the M series SOCs. Having multiple different tiers in each section (M4 low vs. high, M4 Pro low vs. high, M4 Max low vs. high) can be quite confusing and makes comparisons with past generations even more confusing. I am fully aware that every competitor in the space has a naming scheme that is equally bad or far worse, though.
I wish someone started establishing a naming scheme similar to Nissans engine designation. As an example, without knowing anything but a few simple rules, VR38DETT tells you everything about the engine. It is part of the VR line of engines, has a capacity of 3.8L, double overheadcams, electronic fule injection, and twin turbos. Easy to understand at a glance.
Something like that for SOCs would make things far more streamlined. I know that it isn't great for marketing though, so this will remain wishful thinking.
Lastly, would have liked for Ars to include the M1 series as well, though honestly for many, even upgrading from that may be hard to justify. As a previous M1 Max low (24 GPU core variant) owner, I am particularly interested how the M4 Pro high (20 GPU core variant) stacks up in comparison, especially as the lack of 64 bit atomics made some UE5 features hard to reliably implement on the M1 generation, regardless of performance.
I don't know why they are using low/high on the article. It's just tiered like yesteryear's i5 vs i7 for mobile chip; but using i5 vs i7 was highly misleading to consumers because it would lead people to think that there is some great gap between the two chips (when they were the same chip at different clock speeds).
Essentially, these are six very different SOCs sharing only three names. If they aren't clearly separated by Apple, then the media must find some way to communicate that.
I fear though that rather than more clear, the path taken, most noticably, by Nvidia of mixing and matching tiers with completely different silicon will win out, trying to use 4080 branding for both AD102 and AD103 parts certainly didn't harm their long term prospects.
That’s different than 6 very different SoCs as your comment says. Each tier has different capabilities, ie the pro is not a binned version of the max and the base is not a binned version of the pro.
Let's look at Nvidia again. AD102 is used in the RTX 4090 Ti, RTX 4090, RTX 4080 Ti, as well as a less common variant of the RTX 4070 Ti. Each uses the same underlying silicon, just differently binned with certain parts fused off, similar to Apple.
Yet, and this was the point I made, what Apple currently does would be equivalent to Nvidia just calling all of them RTX 4090 Ti (they are the same underlying design after all), with reviewers and customers left to hunt down the specific core counts and differences between them.
And as mentioned, Nvidia tried something even more egregious with the 4080 12Gb, though (this time) were faced with such backlash that they pulled it back. Whether with Apple, Qualcomm, Intel, Nvidia or AMD, every time these practices aren't pointed out by the media, we get closer to a world were a 4080 12Gb will be pushed onto consumers who assume launch day reviews of the "proper" 4080 show equivalent performance.
In that sense, both NVIDIA and Apple are doing the same thing.
Yeah, that it's 1.2 liters too big. RB26DETT me please xD
https://9to5mac.com/2024/08/08/gurman-heres-when-to-expect-m...