Traditionally AMD does NUMA differently from Intel. Would like to see a comparison focusing on NUMA.
Edit: opencl post above says apparently not, it's a different memory architecture.
However, because all cores in a single chiplet share a single L3 cache, doing NUMA-like optimizations will still yield significant benefits.
AFAICT.