Local machines are a sunk cost, so using them is effectively free. Why would you pay a cloud host to run a model that can happily work on your MacBook?
If you're buying a machine specifically so it's capable of running LLMs for you, then the purchase cost is your up-front payment for the inference you'll run.
And between that and electricity costs, cloud has you beat.
So, that’s a decent amount of people who could realize it today!
Apple will be leaning into that further. No other play makes sense.
I see LOTS of text transformation tasks…