I have found out that the mistakes of other models (which I choose first to save money) help me refine the prompt more and more, until I am fed up and pick Opus 4.8 (for example) which magically seems to get it right, but there is a lot of pre-work there...
I wish Fable really was only a minor upgrade so that it wouldn't be missed, but this feels like the difference between having a post-doctoral colleague and educating a student that I have to constantly guide and correct. It's so profound for me that some of the reactions in this thread feel like they come from another reality entirely. Or maybe they just got instantly diverted to Opus, who knows.
More stuff done per dollar or more stuff done for more dollars? Seems to be an important distinction
Fable was definitely better for a variety of tasks, even accounting for using 2X the token rate, like the way it used the tokens faster reduced the wasted tokens, as least for the subset of those who already knew at least some optimizations...?