When it's "that's why they incremented the version by a tenth instead of a half" you know things have really started to slow for the large models.
I.e. it seems we don't get much more than new training run levels of improvement anymore. Which is better than nothing, but a shame compared to the early scaling.