And here I was in bliss with the 32k context increase 3 days ago. 128k context? Absolutely insane. It feels like now the bottle neck in GPT workflows is no longer GPT, but instead its the wallet!
Such an amazing time to be alive.
Such an amazing time to be alive.
128k context is great and all, but how effective are the middle 100,000 tokens? LLMs are known to struggle with remembering stuff that isn't at the start or end of the input. Known as the Lost Middle
EDIT: the above is corrected, it previously erroneously said the non-turbo model was marked as "deprecated", which is a different thing.