Only those who don't care/know about prompt processing speed are buying Macs for LLM inference.
Even useful models like gemma3 27b are hitting 22 t/s on 4bit quants.
You aren't going to be reformatting gigabytes of PDFs or anything, but for a lot of common use cases, those speeds are fine.