I think it's probably more like a year or a year and a half. I don't want to say two years, but it's what I'm actually thinking.
There's simply no replacement for training on more, better tokens, with more parameters. Mythos/Fable was estimated to be closer to 10T parameters than the 800B like GLM 5.2 is.