Stable Diffusion Gets a Major Boost with RTX Acceleration
nvidia.com
nvidia.com
- You have to "pre-bake" each model (create optimized engines) for only a pre-defined set of resolutions
- ~~No LoRA support~~ I guess there is a LoRA support, but you also will need to convert them to TRT
It's been like this for quite a few months now, not sure what's new this time.
This seems likely to be more useful for people with a "production" workflow with particular checkpoint, LoRA, & resolution combo than what I get the impression is the more common hobbyist situation of having a decently diverse collection of checkpoints and LoRA and fairly freely mixing and matching.
On a 2060S, no launch arguments, going from ~4 it/s to ~7.5 it/s.
As someone with a lowly 10gb card sdxl is beyond my reach with a1111 it seems. It functions well enough in comfyui but I can't make anything but garbage with it in automatic. So I'm happy to see 1.5 gets a big boost, I know there's a million of us out there who can't quite squeeze SDXL out so the maturing of the "legacy" versions is a positive note to see.
SDXL reportedly works at below 8GiB in A1111 with --low-vram, and at 8GiB to anything short of 16GiB with --med-vram-sdxl, and 16GiB and up with no options.
> but I can't make anything but garbage with it in automatic.
If you aren't getting out of memory errors with SDXL but the output sucks, its probably not a VRAM problem. The most likely thing, IME, is using the SDXL base model or another checkpoint without a baked VAE but still having A1111 configured to use an SD1.x VAE (or using one with a baked VAE but having A1111 configured to override it.) SDXL needs an SDXL-specific VAE, but A1111 doesn't bundle one, IIRC, you need to download it separately.
There's been a lot happening on SDXL, yes.
> My expectation was that SDXL would probably not see wide adoption for a while because it's hard to make up for the worse training data, is that bearing out?
What worse training data?
Nice! Care to share any projects/advancements/gens you've found particularly compelling if you have anything handy?
> What worse training data?
I don't know this, but I think it's a pretty safe assumption given the state of 2.0/the political climate around this stuff, along with the comparative lower diversity of gens I've seen come out of SDXL. I haven't seen anything to the contrary but I'd love to be wrong.
Where do you read that SDXL has worse training data? They didn't disclose what the model was trained on like the previous models.
> --medvram-sdxl --no-half-vae
In general, these insane speed boosts come at the cost of bleeding edge features.
[1] https://github.com/huggingface/diffusers/blob/28e8d1f6ec82a6...