It's not just TSMC's 3nm process. It's also Apple engineering.
This engineering was common in x86 CPUs by 2013 when AMD introduced Heterogeneous System Architecture which utilized Heterogeneous Uniform Memory Access. The approach has its upsides and downsides, the latter generally being a unified bus tends to have much lower overall bandwidth (even in the Max) and runtime scalability issues. Upsides are more obvious for the types of systems people want APUs for in the first place though so that's usually fine.
The main bit of engineering Apple should be lauded for in the memory department is the gumption to throw the hundreds of GB/s at the mid and high end models.
https://www.realworldtech.com/fusion-llano/3/
In contrast apple actually has everything tied into a single unified space with a single controller that immediately makes all writes visible regardless of where the happen.
They’ve also got enormously more memory bandwidth to play with. M1 Max is close to PS5 in both shader configuration and memory bandwidth.
"Sooo apparently there's this myth that only Apple has "unified memory"...?
Every single modern integrated graphics system works the same way. They even share the MM code in Linux.
I just had a reply guy try to argue otherwise with me even when I told him I wrote the driver"