Stable Diffusion 3 mangles human bodies due to nudity filters
arstechnica.com
arstechnica.com
"We have unique assets and experience which we can bring to synergize with the world-class leading AI efforts that Stability AI has been bringing to market, and we are excited to bring these ------- to the common person on the street." said the Administrator, -------------------------.
"-------!" ---------.
Early testing have users shocked, ---------, and ---------.
"I see the world through new eyes now, all seven of them." reported one user.
"------- ----- - ---------- ---------." reported another.
"We anticipate great profits and ----------" says Stability AI's CFO, shortly before sitting down, sitting down, sitting down, sitting down, sitting down, and sitting down.
---------.
We won't really know for sure unless stability ai comments on it or releases the training data but it is the most likely explanation.
Someone will likely train this model on porn anyway in the next few months so we'll see if the finetuned models work better.
Except this time, it will be even worse. SD2's lack of good human generation to correct its faults could be fixed with finetuning, but SD3's noncommmercial-unless-you-pay-us license will disincentivize any community finetuning.
But for the prompts I actually tried, the speed and prompt adherence was a significant step up. The overall quality does not compare at all to SDXL fine-tuned.
I think overall people should appreciate their improvements to the architecture. Although the deficits are bizarre and point to a huge hole in the market for new companies to get involved in pre-training models.
But there will be fine-tuned on Civit AI.
Also, people kept pointing out how good PixArt is.
Due to the noncommercial licensing, there will not be nearly as much finetuning as with SDXL.
And then Lora’s are small enough to do without large resources
Loras are easy to train on consumer hardware, but a fine-tune like Pony takes significantly more resources and a lot of work in terms of image tagging, curation, etc.
One side question I haven't seen much (intelligent) discussion of: to what extent is the intensity of AI critiques disingenuous in origin?
Unfortunately, disingenuity discussions become instantly meaningless when they devolve into politics. Better to just assume that all politicians are disingenuous in everything and move on.
I would say there's a category of unpardonable voices, such as competition-suppressing established companies (and wishful ones like musk.ai) and dick-tators targeting AI as a proxy for Western economies. However, I'd pardon anxious content creators, the same way I'd pardon Lyft drivers critiquing self-driving cars (and cab drivers critiquing Lyft 10 years ago.)
The Goldilocks region is elusive. For example, it's unclear if SD3's failures are actually due to NSFW filtering. I suppose the disingenuity test could be based on the speaker's technical credentials, but that would probably dismiss too much speech.