(The silly story they have on the site doesn't really score any points either, it reminds me of RIAA et al)
(The silly story they have on the site doesn't really score any points either, it reminds me of RIAA et al)
It has more in common with using an AI language model as a "bullshit generator" than it has with human attribution. Which isn't a coincidence, since nearest neighbour similarity searches ARE a type of machine learning, just a much simpler one than SD.
Anti AI people who are upset about attribution, should learn the technology and try to create an actual attribution model: one that has a notion of causality, which could say who influenced who.
Causal model requires causal assumptions, but you can probably get far with the simple assumption that "works from the future do not influence works from the past".
It still finds ostensible "source images" for the art. So, yes, it's clear this service is pretty much bogus.
* https://en.wikipedia.org/wiki/Julie_Wilhelmine_Hagen-Schwarz
I don't think such a solution can work just based on the resulting image though. It probably needs to input prompt to have any chance of working.
Wouldn't the attribution be basically all the dataset with various attribution probabilities? It's kind of like a reverse neural network.
More so how are these image generators any different from text generators like ChatGTP? I feel like if first tool out from these AI gates was a bot that wrote good-enough-to-use code, no one would have batted an eye. Everyone would just go "yeah we told you that those pesky programmers would eventually automate themselves out of jobs", but since it is rendering pictures which most of the populace can appreciate and it is a skill that is easy enough for literally any child to pick up, but requires a lot of dedication to master all of a sudden this is theft and should be illegal.
I really hope this lawsuit or whatever doesn't go anywhere, because it will not change anything. The genie is already out of the bottle as far as image generators are concerned - however it means that we the people won't get whatever comes next. Whatever comes next will be tightly held by big corporations and they alone will reap the benefits, whatever they may be.
I've spotted this pattern a couple of times and this sort of circular reasoning seems concerning. Whenever one of (Stable diffusion, Copilot, ChatGPT) comes up in a discussion, their legitimacy seems to be swiftly justified by existence of the other two, even though they're all uniquely problematic in how they wash away attribution and licensing.
This is not circular reasoning, more observation of what people value. Since everyone can "Google and gather information" ChatGPT isn't valued enough, but since most people don't know how to draw DALL-E and Stable Diffusion are seen as industry destroyers.
The difference is I don't see book authors and news writers trying to sue OpenAI for "stealing" their articles and not providing attribution, or creating websites that try to divine which specific book chapters or news articles ChatGPT used to generate any particular response (as if that's at all representative of how GPT works).
By comparison, a “dog vs. cat” classifier that has 100% accuracy on the dog/cat task will, nonetheless, tell you that a slice of pizza is a dog... or a cat.
(You could possibly interpret the results as “if SD had generated this image, it would have drawn most from these sources in doing so”.)
Note that just because an image is similar (to human eyes) doesn't mean that it played a more significant role than a seemingly more dissimilar image. It could even return similar images that SD wasn't trained on at all. Even conditioned on providing an SD-generated image, it fails.
(Something doing what it claims to do, as opposed to naive image similarity, would actually be pretty cool and useful.)