Kernel scheduling is NUMA aware and will localize workloads.
Threads will mostly have their RAM on the sticks local to their node. The core the thread is delegated to is also more likely to be the core local to the disk or NIC being used for IO.
This is at least my experience, though I am no expert.