The most interesting aspect of this research is that Google published it. They have a huge competitor in this area, it would make sense to stop publishing unless they think that OpenAI already does this.
Stable diffusion was possible because diffusion model architecture and then doing it in latent space instead of image space and then upscaling.
Of course someone still needed to train it, but they did it when it was theoretically possible.