I came across this article the other day for the first time and started to read it. How much still applies to today?
P.S. What’s I dislike most about the article, it fails to explain why L3 cache is 10-20 times slower than L1 cache, while they both made from SRAM.
Why is it? Is it because L3 is usually shared?
The best explanation I saw is this: https://fgiesen.wordpress.com/2016/08/07/why-do-cpus-have-mu...
All the timings are a little faster, the caches a little bigger, and the buses a little wider, but it's still basically the same stuff with different names.