Yep. They could've had socketed RAM but they wouldn't have gotten the same bandwidth and that's really important for running things like LLMs.
and games, having really speedy/low latency memory helps a lot with being competitive to some mid range dedicated GPUs
Unified RAM/VRAM is very nice for running LLMs locally, since you can get wayyyy more RAM than you typically can get VRAM on discrete GPUs. 128GB VRAM on discrete GPUs is 4x5090s — aka $8k just on GPU spend alone. This is $2k and it includes the CPU!
Of course, it'll be somewhat slower than a discrete GPU setup, but at a quarter of the cost, that's a reasonable tradeoff for most people I'd think. It should run Llama 3.1 70b (or various finetunes/LoRAs) quite easily, even with reasonably long context.