Prior to LLaMA + llama.cpp you could maybe run a large language model locally... if you had the right GPU rig, and if you really knew what you were doing, and were willing to put in a lot of effort to find and figure out how to run a model.
My hunch is that the ability to run on a M1/M2 MacBook is going to open this up to a lot more people.
(I'm exposing my bias here as a M2 Mac owner.)
I think the race is now on to be the first organization to release a good instruction-tuned model that can run on personal hardware.