Per the code, the technique is based off of DeepFloyd-IF, which is not as easy to run as a Stable Diffusion variant.
Per the code, the technique is based off of DeepFloyd-IF, which is not as easy to run as a Stable Diffusion variant.
I haven't dug in yet, but it _should_ be possible to use their ideas in other diffusion networks? It may be a non-trivial change to the code provided though. Happy to be corrected of course.
> Our method uses DeepFloyd IF, a pixel-based diffusion model. We do not use Stable Diffusion because latent diffusion models cause artifacts in illusions (see our paper for more details).
The ecosystem around Stable Diffusion in general is so massive.
Is there a good repository anywhere or is it just "wade through twitter"?
https://www.latent.space/p/sep-2023
https://github.com/swyxio/ai-notes/blob/main/Monthly%20Notes...