At this price points (including the reduction of usage on the Pro 200 and Pro 100 subscriptions), more people will be looking buying their own hardware and running with open weight AIs. $6000 per year is the threshold
It's a bit more complicated than that though. Dollars per token is not the right metric to track.
Local models don't hold a candle to SOTA models unfortunately.
Does it matter when I burn through all my Fable subscription usage in about a day? Astra is much better on efficiency, but doesn’t last a week either. With the new subscription limits, it’ll end up being like Fable.
You should be using Opus 5.5, it's just as good if not better than Fable, and it sips tokens.