This basically creates a bottleneck at the oldest/cheapest Apple Silicon machines, which are already crippled for context prefill.
This basically creates a bottleneck at the oldest/cheapest Apple Silicon machines, which are already crippled for context prefill.
But honestly, obsoleting a huge number of otherwise great Apple Silicon machines is something Apple would moment consider a major "pro" of building a compelling local AI stack.
With how much speculation around the difficult time Apple has had getting people to upgrade from M1, I'm sure they'd jump at such an opportunity.
- Please buy our new Macbook pro M5 that gives you 20 tokens/s on local 80B LLM
next year - Please buy our new Macbook pro M6 that gives you 25 tokens/s on local 80B LLM
milking product revenue in perpetuity by offering meaningful marginal improvements, while keeping same architecture will be the golden goose for Apple
+plus if it allows to segment market by wallet size into poor/middle/rich classes, thats even better