If we wait for the next models, we will never test anything because there will always be another model. Like the Ai Scotsman:
> "Nay, laddie, that’s no’ the real AI Scotsman! He’s grander still! More powerful! Just wait for the next model!"
I think he’s being sarcastic
Also throw in GLM 5.2 for good measure
That will be in the Part 2 article.
I worry that GPT 5.6 will be heavily restricted and have the same feature to fallback to another model like Claude fable 5 does all too often. That fallback shenanigans mess up actual benchmarks and I don't like it.
Well, I'm probably not on the list of special people who will get to see GPT-5.6 Terra.