AI Art Panic
opguides.info
opguides.info
What’s likely to happen is people are going to invent whole new media that will use this. Incredibly elaborate interactive and personalized experiences will be made possible by this. One person will be able to deliver a huge amount of value, and they will be rewarded accordingly.
Think back to the 1980s. Programmers were in very short supply, and the work they were doing was less accessible and more difficult than it was today. IDEs, version control, better languages and frameworks, easy access to tools to learn to code, Stack Overflow, etc. made programming much more accessible and engineers became much more productive. The number of programmers out there is orders of magnitudes greater than three decades ago. Yet real salaries have still grown tremendously. The value per capita has grown even faster than those salaries because they’re so productive.
The same thing is going to happen here. Artists will build unimaginably huge, complex projects (on the scale of Google/Facebook/Amazon), and the world will gobble them up.
Suddenly, he’s got a cheap option for getting some basic art for a simple turn based game, card game or similar.
If a game of his takes off? He’ll have to hire artists to keep up with the increased workload.
AI will move barriers of entry lower, not higher.
Steam had almost 9000 new games added to it last year, and the average revenue of each fell by almost half a couple of years ago
If these tools had been available when I was a kid, my tile based RPG could have had the graphics I always wanted it to have. Instead, it never got anything more than map tiles (primitive ones at that). I had no talent or skills for asset creation.
Nobody played that game other than me, and (once) a couple other people I knew. That was fine, I made it for me, not others. But I would really have been happy to have access to assets.
I don't know such a law, but it's simply false. People have 2 finite budgets they can allocate to entertainment, time and money. In many ways money is not the limiting factor anymore since people have Steam lists with hundreds of unplayed games and the abundance of Free2Play. The main factor it not to have a saleable product, it is to manage to get in in front of enough eyes, with all marketplaces already being saturated and ads being useless unless you have a very high budget.
I'm a bit worried that it will devaluate indie productions even further because of an over-abundance of AI generated content, further elevating industry production that can afford really competent artists that you can't even start to replace.
Namely, it is more than time alone, it is "that right time" and there is much, much less of it.
> But as the cost of something falls, people consume more of it. That’s a law in economics.
Which law is that? When the tulipmania crash happened, did people start consuming more tulip bulbs?In any case, it is a bizarre notion with perhaps limited applicability in the real world.
Tell me about your thoughts on the marginal utility of child pornography.
Need to be careful with those universals.
“conversely, as the price of a good decreases (↓), quantity demanded will increase (↑)”
But not so the effort needed to produce "unique art".
When everybody can push a button to produce "art" then how do you make the results of your button-pushing unique?
The value of art is in its uniqueness, and originality.
Using AI to generate art based on input training models is not that different from photography. Each generated art-piece is like a photograph taken from different viewpoint. And when everybody can use the same programs to generate their "art" the results are bound to be more or less non-unique, non-original.
The scary part for artists may be that now programmers can create new original generated works of art. You can be an artist without knowing how to draw.
Then again we could have a conventional artist with a unique style who would train an AI model to only generate versions of their existing paintings. Laws may need to be adapted to make it illegal to generate art based on someone else's existing art. Lawsuits will follow.
https://en.wikipedia.org/wiki/Jevons_paradox
In this case you can think of Artists as the "resource" being used.
Have you talked to any artists? Because in my experience they universally hate AI generated art. Not just a little. I think I've seen more artists who were open to NFT's at this point than AI. Art sharing websites are getting filled to the brim with posts from new accounts sharing AI art which sometimes very directly takes clearly identifiable features and styles from other artists.
AI generated art is beckoning a possible dark age for visual arts. There are many lost arts already, if AI trivializes the costs and efforts why bother honing it. Then we find ourselves in a creative void as the only art generated will be based on art that of which has come before.
In my experience of visiting art museums, some of the classical and renaissance art is incredible. There is clearly some profound idea behind the piece, and the skill is the vehicle of that idea. A huge amount of art, though, is simply commissioned portraits of some wealthy businessman or forgettable Duke. The purpose of these paintings is simply to represent reality. There is no idea behind it.
For the artists painting them, the camera must have been devastating. It meant that years of training in studying how to properly draw an eye and how to get the colors inside a shadow right were at risk of being entirely devalued. But the camera also laid bare the fact that art is more than just skill. I imagine a lot of artists who were coasting by on skill alone hated the camera.
We’re at that point now with digital art. The cost of skill in actually drawing a thing is now zero. Anyone can do it. The question is, do they have anything to say?
Excellent point. It's up to us to level up to the opportunity.
It's happening very quickly and the outcomes are difficult to predict, but it's clear it's a big deal. We either fear the unknown or are excited by it. I feel most of the controversy is in fact the robots-will-make-our-life-comfortable / robots-will-take-our-jobs discussion. The real problem is elsewhere, not the tech itself.
Photography is if course a good comparison, but don't forget 3d CGI. I can place things exactly how I want them and have the computer do ALL the drawing for me? Outrageous! It certainly removed a lot of need for menial manual work (and even filming), but has also opened the opportunity for creating entirely new things that before would just not be practical, or enjoyable to make. And there's no doubt now that a lot of skill is required to make things look good. It can seem like it's easy to get great looking stuff with AI but it's not. What's easy is to churn out tons of samey derivative crap that will age very badly as soon as we get used to it, like with early 3d.
Yes, good point but let’s look at what a a boring mess CGI movies have become. Movie productions used to employ a lot of artists and looking back you actually can tell. For example the 1982 movie ‘The Thing’. All those effects in the movie are realistic, are not CGI and have become now a lost art.
Photography can capture reality far better than any artist could achieve even with a lifetime of practice, and today it can do so with nearly zero cost--take out your phone, hold down the camera button, snap a picture. These advances hardly put an end to drawing and painting as skills.
It turned out that there's value in art beyond verisimilitude, and if anything the advent of photography encouraged a resurgence of exploration into what art really is: a movement away from pure verisimilitude in search of feeling.
This is what I expect to happen with AI art today. Yes, the AI models can produce art that looks awesome and wonderful now, but it's frankly already starting to wear thin. Art exists to fulfill humanity's need to experience something new, and AI alone simply cannot keep up with the meta game.
There will need to be human artists guiding the AI to produce something valuable, and there will continue to be human artists working by hand to create the novel works that the AI could never have rendered.
Yes but who’d want their art to be fee back in trainig the AI to achieve what they have done with a click of the button?
> This is what I expect to happen with AI art today. Yes, the AI models can produce art that looks awesome and wonderful now, but it's frankly already starting to wear thin. Art exists to fulfill humanity's need to experience something new, and AI alone simply cannot keep up with the meta game.
But the essence of artwork can actually be copied away too and in the process squander the artists livelyhood. Yes, the argument that the artist can themselves take advantage of that and that is true for some computer literate artists but what about the others? Does the future of art necessarily have to tie art to computers and AI?
No, because the essence of art is meaning, and AI cannot deliver that without a human agency behind it.
AI will do to art what technology does to every field: it will eliminate the need for shovelware-style jobs, leaving only the truly creative jobs behind.
It will most likely resemble what happened to web developers when SquareSpace and WordPress became good enough for untrained people to pick up and build with. If all you knew how to do was write enough HTML and CSS to make a website that looked "good enough", your job was in danger. But if you were a good designer or a good engineer, you transitioned to a more specialized role, working for the many orgs who needed something beyond the capabilities of the tech.
I expect the same will happen in art. To use book covers as an example: the job of illustrating covers for cheap romance novels will certainly be replaced, because they all look the same already. But publishers will still need humans (either driving the AI or painting by hand) for the books they hope will be best sellers, because it takes a real artist to know how to capture the spirit of a book in a cover image.
I see no reason why another AI eventually couldn't come up with prompts from human interests found from algorithms. There's exactly no reason a human needs to be involved at all.
The main issue is that anyone can write a simple prompt and get 85% of the way to an image that previously required a decade of experience to create.
But that last 15% is where real artists differentiate themselves. It’s the last 15% that can’t be faked if you don’t have a sense of artistic direction and discretion.
I strongly disagree with “why bother honing it” as well. If anything, being able to get to the “honing” stage so quickly is what makes these tools so powerful. I’m sure if Gordon Ramsey had to butcher his own chicken, grind his own grain for flour, and churn his cream into butter, he wouldn’t spend as much time on the finishing and plating of a chicken sandwich.
With the limited subset of artists I have worked with (traditional modern / contemporary painters) they have been very inspired by the creative possibilites of these models and in one case see it as a great opportunity to scale their work to new types of media such as video.
It would have been practically impossible for one of the artists I am working with to draw each frame by hand but now she can generate music videos in her style by tweaking and working with Stable Diffusion and a set of reference images.
I don't think the invention of a camera displaced many artists, but I'm quite sure that they were just as angry, probably saying how it doesn't have a soul and such. But in the end it improved the life of humans to an unimaginable levels.
I think same if not more can be or will be achieved with AI generated art. And I don't think it's displace many artists jobs.
Again, the coming of AI art feels like the coming of camera must have felt.
AI threatens to make artists obsolete by via the direct usage and synthesis of their work along with the work of their peers. Used without their permission.
As someone that works in generative modeling I think it is important to note that we are hand selecting results and not showing the average ones. There's been such a hype machine that if we show our average results that we'll get killed in review (publishing in a hyped space unfortunately requires a lot of salesmanship). Not a thing I'm particularly happy about, but I think this is important to note because if you're just looking at the images on Twitter, Blogs, and other social media then you're seeing the top 1% results. Images that take quite a bit of time and effort to produce. Likely not as long as an actual artist to make them, but still a lot of skill is involved.
I do agree that I think it'll be a big boon for artists. I've even been using these to get better at drawing. The results are bad even with prompt engineering, but hey, my human mind can fill in the gaps and it is good for expanding creativity. We'll get better at these models for sure but I'm not quite convinced it'll kill a lot of art jobs. But maybe I don't know what those jobs are that are being killed. But then again, I've even seen artists do much better at these models than me even though I understand the code and math going on. There's also cool things like walking through the latent spaces, but you won't be able to do that with the hugging face code. But that is art that you won't be able to do elsewhere. It is entirely new.
https://huggingface.co/spaces/akhaliq/Ghibli-Diffusion
https://huggingface.co/lambdalabs/sd-image-variations-diffus...
https://huggingface.co/spaces/shi-labs/Versatile-Diffusion
https://huggingface.co/spaces/akhaliq/openjourney
https://huggingface.co/spaces/anzorq/finetuned_diffusion
I'd still like to get some good cartoony results with SD, but it's clearly not possible yet.
I even tried to use DreamBooth to train on my own inputs and the results are not really usable, and that was with SD1.4 and AnythingV3 as base models.
But suggestions I'll give is that you need to add detail to the prompts. SD also has negative prompts which really help. So use a lot of words if you can (twitter accounts will help you find magic phrases like "bismuth neon chrome copper patina verdigris"[0]). But you may even see here that the magic words with SD aren't as simple as with DALLE
I do think you're understanding, though, what my post is about and why I think a lot of people are talking from a naive perspective. If you only show the wins you have some extreme biases.
[0] https://twitter.com/sureailabs/status/1596211001120153601
With regard to your main point, I generated more than a thousand images trough various models, and yes good results such as these posted on social media are very rare. And people definitely spend lots of time to adjust things with inpainting and other tricks, in a way that I would find infuriating. For the effort it takes I would rather just redraw it. It depends on the styles; with brushstrokes you can hide a lot, with cartoon lineart is crucial so fixing requires more work.
?
Has it?
Look, you can argue about whether it’s morally right or not to use models that are fine tuned explicitly to copy the style of someone else, trained on their art without their consent, to make a model that can generate images very similar to the training images.
You can argue about technically of that’s copying, or if lossy compression is copying.
…but legal and moral are different things, and right now, as far as I’m aware:
- it’s only legal because there are no laws specifically making it illegal currently.
- there are active (eg. Copilot) cases challenging this to set a precedent.
- it’s sufficiently ambiguous having a model that anyone can type “a naked picture of a 12 year old” in and get exactly that as output, that stability has nerfed the most recent mode release they’ve done.
- there is a reasonably obvious similarity to other fields where a thing is itself not illegal or bad, but it can enable people to do illegal or bad things, and therefore access, ownership and usage of said things (eg. Hand guns) is heavily legislated.
I think “this has already been decided to be legal” is a blatantly false assumption.
Also, this looks to me like a situation where generating art is suddenly so cheap and easy it doesn't matter what the courts decide. People can and will ignore the law, because it is trivial to generate new pictures and it is cost ineffective to enforce any restrictions. We're going to see this tech take off. Who knew that art would be the next thing software took out?
People look at art and make art in that style all the time, now a machine exists that can do that. Why is that unethical? Because they didn't consent for the machine to look at it/learn from it, but they did for humans? I don't think this argument will be able to hold the wave of change that's coming from this new capability. They'd be better off long term learning how to use it.
Nobody creates purely original things in a vacuum, machines won't either.
My brother is an artist by the way who does a lot of work with AI, 3D printing and internet. Art will always survive and adapt. If memory serves it took decades before photography was socially accepted among the art world.
But when I think of the value and interestingness of art, there's a lot of intention and meaning in choosing what to draw, the form, etc.
For example, look at The Death of Socrates and then compare it with the three other paintings of the same scene here: https://en.m.wikipedia.org/wiki/The_Death_of_Socrates.
Even from a first glance, the one by David is clearly superior. The scene is so much more striking and interesting to look at. You can see the emotions of the characters, and overall the composition underscores the significance of the philosopher's death, especially in the context of the enlightenment/romantic period. (This is my opinion as a pleb; not an art expert.) But what I do understand, as a computer person, is that these concepts are still beyond what a model can encode in an image as of today.
So although humans are no longer superior in the mechanics of producing images, I think in the higher-level/psychological aspects of "art", there's room for humans, at least for now.
And although I'm just an dumb anon on the internet spewing these ideas, I know I sound harsh, but I think I'm correct. Because the reality is that now, everybody's downloaded the SDv1.4 weights onto their hard drives, and the cat's out of the bag permanently.
I have never used any of those services before. I have never considered using one of the services before because it was always outside my price range.
Now that I have tried stable diffusion and have photo bashed and rendered some concept art for each of the main characters in my novel, I now want to commission an artist to create the 30 or so needed training images so I can ask stable diffusion to spit out my own character in various poses and expressions.
At minimum, that will require a human to render a model sheet of the character from front and back and side and 3/4 and above and below, as well as the emotions on the basic emotions wheel.
If I want the character to be able to wear different outfits, then I will also need to pay for renderings of that character wearing that clothing, all in service of trying to train stable diffusion to be able to remix that character into future images.
Let's also say that I do not have a killer graphics card to be able to train images into a model, luckily, for another $100 fee, the artist will use their existing graphics card to spit out an embedding or a hypernetwork or a VAE or whatever it is that you can use to add custom training to a model and send me that as well as the original set of input photos.
After all of that, I can generate the photos I want...but I will then have to slightly tweak each image so that it has human authorship, even if it's just removing noise and fixing the cursed loops that happen on limbs at times.
In short, I am considering something that is at least $200 for the crappiest cheapest artist out there, multiply that by my six or so main characters, and that is money that I am genuinely considering spending that I would not have even dreamed of entertaining for a moment.
The proof is in the pudding as to whether this thought process will be happening for other stable diffusion users who are able to get images they like, but do not have good rendering skills on their own, weather they too are willing to pay for this or not.
If so, there will instantly become a new type of artist job available, that of the AI art trainer artist.
At the very least, there will be AI art cleanup artists that remove the noise and so-called cursed elements of ai art when used in the concept art stage.
Currently, we (mostly) still need people to make the tapes. That probably won’t be true forever, but it is for now.
People on HN have very strong options about this stuff without looking at any of the relevant case law.
And why should that be a problem? This is a King Canute and the tide scenario. We may as well bow to reality and admit that it is ok.
Now it’s so easy to share music, it’s meaningless to try to stop people doing it.
Apply 20 years of law cases and punishment for random people and now…
…everyone pays to stream their music.
Right? Wrong? Eh.
I’m just saying, you are kidding yourself if you think that the Powers That Be will just let people decide copyright isnt a thing anymore because of (insert reason here).
Once there is money involved, there will be court cases, and you know, I’ll be shocked if a combination of “needs bigger GPUs to run” and “legal issues” don’t cause these sorts of models to be locked away behind cloud APIs in the future.
It is what it is. Enjoy it while you can; some things are quite predictable, and:
“Law takes a while to catch up with new technology, but it eventually does, and when it does it favours the status quo”
Is one of those things.
Not always; eg. Uber, but… predictably often.
I will point out that the usual copyright maximalists have been pretty silent on the issue of AI art. The biggest opposition to AI is coming from the Free Software community - i.e. the people who want to abolish artists' ownership over their work outright.
I'm surprised here; I thought it only cost like $600k to train SDv1.4. Plenty of wall street folks and programmers have that much money to burn.
You don't need to pay to listen to music if you don't want to. The corpus of good music on YouTube for free is probably bigger than what you can listen to in a lifetime. If it isn't already it will be in time.
Napster's model won that war. Effortlessly. If you're paying for your music, you are paying on your terms based on the value you think is being provided to you. It isn't a legal framework making you do it.
My point is that Napster-to-SD is not a fair comparison. Napster didn't eliminate or replace human artists, while SD certainly did. Therefore, it's not fair to assume the government can regulate themselves out of this one like they did with Napster. Because even though Napster enabled widespread piracy, the human musicians still had leverage in that they were needed to create new music.
So my answer to the question posed by parent comment of "will SD play out the same way that Napster did?" is "no" because there are fundamentally different economics at play.
Obviously, if a good generative model comes out, the music industry will be in a similar boat to the art industry right now. Google was working on it in 2017 (https://magenta.tensorflow.org/performance-rnn) but I don't know if they've made any progress since.
I'm going to need to see a source on ripping the same color palette as being illegal.
I bet Pantone believes they can copyright a color palette, considering Adobe just pulled their colors out of Photoshop.
The other is amusing in light of this case: https://www.supremecourt.gov/DocketPDF/21/21-869/213328/2022...
This is case with anything…anything not illegal is legal, at least in the US…
I think you're dancing around with words here. Let me be super specific:
There's no law specifically saying I can't use a jelly coated toasted to beat someone to death... but there are existing laws that cover 'beating someone to death'.
If you beat someone to death with a jelly coated toaster, you might argue for some obscure reason, your actions are not covered by the existing legal framework around beating people to death, but mostly likely you will not be protected by the claim you 'didn't know it was illegal' to do that.
...because the courts have not ruled that beating people to death with jelly coated toasters is legal.
ie.
a) There is no specific legal precedent or law around something but other laws related to it exist
and
b) Something has been determined to be legal by some precedent / law / whatever
Are not equivalent.
Regardless of what people want to believe, or have opinions one way or another, the assertion, made in the original article, that (b) was true, is not true.
Only (a) is true, and that is, legally, a much weaker statement.
A new law that is passed doesn’t have retroactive power
A court ruling that the thing you’ve been doing for a year has been breaking an existing law can absolutely punish you for it. But it’s unlikely to punish random individuals doing things on a non noteworthy scale when a clear understanding of the law in a new context has not been established.
At least, at this time, with current interpretation- "the Supreme Court has explained that people must have notice of the possible criminal penalties for their actions at the time they act" [2] (See also Weaver v Graham[3])
This might be subject to change.
[1]https://constitution.congress.gov/browse/essay/artI-S9-C3-3-...
[2]https://constitution.congress.gov/browse/essay/artI-S9-C3-3-...
[3]http://cdn.loc.gov/service/ll/usrep/usrep450/usrep450024/usr...
(It’s legal the same way Google Image Search is.)
I could very easily see it being argued that “computer generated images” are in the exact same marketplace as “images” so already that case wouldn’t apply and new legal reasoning would be needed.
This is why we do heath inspections etc, it’s all about shifting the risk vs reward calculation.
In this case, the riskiest issue isn’t copyright, it’s that it can generate NSFW/CSAM, which governments and payment processors both get really upset about.
It’s part of the basic structures where the legal team has an advisory role.
That's obvious hyperbole and not a useful, or even factual, rebuttal.
Much of this is simple ignorance and many thing aren’t particularly relevant. 14 states still had sodomy laws in 2003 when the Supreme Court reversed its stance and declared them unconstitutional. At this point there are hundreds of years of crap at the federal, state, and local level much of which changes based on where you happen to be.
What percentage of the US laws have you actually read?
One, that there is always a company out there somewhere acting criminally. If that's the intended meaning, it is a factual statement, but doesn't not carry the original implication that companies do not try to avoid breaking the law. For instance, the fact that there is always a human out their somewhere committing crime doesn't not mean most humans do not actively avoid such.
Two, that any given company breaks the law often. This carries the implication that companies do not worry about breaking the law, but is also factually incorrect.
Same with copilot; people copy shit from GitHub and SO all the time without mentioning copyrights and a lot of code people cough up is just ‘stolen’ from someone without remembering who/where it was; what’s the difference?
Lots of bands (nearly all?) learn their craft by playing covers. At some point the artists in the group start to find their voice.
I suspect AI never will.
They are both end 50s, so I rather doubt it.
Court cases / legal clarification may outlaw Stable Diffusion or DALL-E for having been trained without the consent of the original artist on their copyrighted artworks. But if this technology proves valuable and viable, the likes of Disney, WPP, and Omnicom Group will pay a few dozen artists to create enough work to seed an engine that can generate 50 million wholly-owned lookalikes.
Copyright protection will ultimately shield the megacorps from competition by the common folk, not artists from competition by corporations.
No, it's only legal because nobody has litigated a case all the way to the Supreme Court. We all thought that APIs weren't copyrightable, and then, suddenly they were (or worse, were "assumed to be" but not explicitly ruled upon).
One might argue that any unlicensed use of copyrighted material is misuse (unless the courts have ruled it is fair use.)
Only the copyright holder may bring action against a potential misuse. Typically, individuals tend not to have deep enough pockets to hire the lawyers to bring the action. And during discovery it may be found that the model was trained on works-for-hire made by humans replicating the original style. Or appeals may continue on at more legal cost, without guarantee of recompense.
Most will decide it’s just not worth it. Until a company similar Getty buys up all the copyrighted material and, in addition to sensible suits, brings spurious legal action against individuals and turns the tables.
Unfortunately art is a hard business to make a lot of money in and the vast majority of the people unhappy about this do not have the money to mount any kind of legal challenge to the corporations developing these databases, so we are basically fucked.
Which is in some ways nothing new, the average Internet user has no concept of creator's rights and will blissfully assume that if something's not behind a paywall, it's free for any use. It feels even shittier and worse when this is being done on an industrial scale by these AIs, though.
Open source models will likely be dominant (this was how Stable Diffusion got popular, it's not clear their new neutered 'safe' model will be as popular or a successful business).
This is the new reality even if the bigco gets sued out of business. Artists need to accept that reality eventually. Just like all "panics".
Fortunately AI generators are not a direct replacement for artists, I've seen it used as part of the design process, but that still takes talent. Maybe it will get better but it's not a one-stop shop for business use-cases. More likely it will be AI+artist, not AI replacing artist.
A real open-sourced model trained on images that were explicitly licensed for such uses would be interesting. It would cost money but it would also not be hiding a ton of its actual cost by training it on images whose creator never imagined that "being dumped into a vast dataset" would be a thing that would happen, as well as images that were fine in their original context as "fair use" but got sucked in by the web crawlers feeding the dataset. You want to train an image-generating machine on my work? Ask my fucking permission and if I say "pay me" then you either negotiate a price we can both live with, or you do not put it in your dataset.
I do not accept your reality. Sue 'em all until you've gotta train your own damn datasets, or buy ones that have worked out the proper licensing. If Photoshop can be made to refuse to scan currency, then these things can be made to be fair to the artists whose shoulders they rest upon.
In neither case is it legal to use outputs of an AI that infringe an already-existing copyrighted work. This means anyone using it to generate production-ready artwork is exposing themselves to potential legal liability if their AI winds up regurgitating training set data. This is what I worry about way more than just "are the model weights infringing copyright".
As for morality:
- Absolutely none of the current generative art systems are trained from scratch on ethically-sourced datasets. The AI companies just assume that because they need ungodly amounts of training set data, that they are morally entitled to get it, because they spend more time staring into the eyes of a basilisk[1] than worrying about the world they already live in.
- Multiple companies getting into generative art have gone above and beyond in giving human artists the middle finger. DeviantArt decided to blow all their good-will from defending against NFT nonsense by making their AI training program opt-out, which is NOT HOW CONSENT WORKS. Mimic[2] and Dreambooth are basically artistic impersonation tools that have specific moral implications beyond the general practice of AI art.
Whether or not this becomes actual law is... well, I'll put it to you this way. Stability AI's Dance Diffusion is actually trained on an ethically-sourced, public-domain dataset. Why? Because they're afraid of being sued by the RIAA. Despite the name, copyright maximalists are less about maximizing the rights of artists and more about building moats around large publishers. And generative art does not threaten[3] those publishers, so it will not be banned.
[0] Yes, the same one that mandated upload filters on video sites.
[1] Roko's Basilisk posits the idea of a superintelligent AI - one that can eat everyone's brains and emulate them perfectly - constructing a perfect hell for people who didn't build it as a way to threaten those people in the past to build it.
Longtermists can burn in computer-generated hell. Time discounts exist for a reason.
[2] https://illustmimic.com/en/
[3] Specifically: even in a world where AI completely outperforms humans on all artistic endeavors, publishers will just buy the AI companies.
I have investigated this a few months ago while researching a radio show segment on AI art. In German law, the relevant statutes appear to be section 23.1 and section 44b of the copyright code. The original text is at https://www.gesetze-im-internet.de/urhg/__23.html https://www.gesetze-im-internet.de/urhg/__44b.html and below is a Google Translate result that I have proof-read.
> Section 23 (Adaptations and rearrangements): "(1) Adaptations or other rearrangements of a work, especially a melody, may only be published or used with the consent of the author. If the newly created work is at a sufficient distance from the work used, it does not constitute an adaptation or rearrangement within the meaning of sentence 1."
This amounts to the same as what TFA says about derivative works in general. If it's too close to the original, it may be infringing copyright. Gauging the boundary between "too close" and "not too close" is the bread and butter of copyright courts.
> Section 44b (Text and data mining): "(1) Text and data mining is the automated analysis of one or more digital or digitized works in order to obtain information, in particular about patterns, trends and correlations. (2) Duplications of legally accessible works for text and data mining are permitted. The copies are to be deleted when they are no longer required for text and data mining. (3) Uses according to paragraph 2 sentence 1 are only permitted if the right holder has not reserved them. A reservation of use for works accessible online is only effective if it is in machine-readable form."
Again, IANAL, but in my opinion this covers neural networks. "Obtaining information about patterns, trends and correlations" is as close to a literal description of the purpose and function of artificial neural networks as you're going to get in legalese. Paragraph 2 just means that you have to delete your data-mined stash if you ever deem your model complete (but it should not be too hard to argue that ML training is always an ongoing process and thus the mined data will always be required). Regarding Paragraph 3, if I'm not mistaken, the "machine-readable form" part of paragraph 3 is very very very specifically aimed at robots.txt. That is, if you allow your art to appear in search results (which most artists want), then it's also fair game for data mining.
Once again, IANAL, and this is only German law. But I think it's very easy to make a case here that this existing code of law covers AI training as well, as long as measures are taken to ensure that source images cannot be reproduced with any sort of fidelity (and obviously that's a big "if").
I had to say it.
What does this mean?
> txt2img prompt crafting is a bit of an art in its own right, as getting the model to spit out what you want isn’t always trivial.
I don't find the output images from a lot of "AI" art to be artistically interesting because an "AI" made them. The output might be novel, or even kind of surprising, but still just the output of code. Someone might say the code of an "AI" art machine could qualify as art. I often find the txt2img prompts artistic in a way that the output images lack, because they are representative of human imagination, often like surreal poetry, or like musique concrete, cutting and pasting segments around on a tape. This is not to deny that the output of "AI" art is impressive. It is, frankly. But as a matter of machine, computational, output rather than human artistic merit.
It comes down to a simple thing: if you showed me two images and you told me a human made one, and "AI" made the other, I will always find the former artistic in a way that I would never find the latter.
This is also not to touch on the licensing issues that arise from these tools. That's its own knot to untangle.
It doesn't matter anymore that you emotionally value, or want to emotionally value, human-made art more than AI-made art, if you can't distinguish the two except based on other people's representations of how it was made.
Unless you watched a (non-augmented) human make a piece of art, you have no assurance that a human didn't touch up an AI generated art piece or simply curate a collection of fully AI-generated art.
ETA: I'm not trying to put down recognizable human-created art that's already been created, especially historically notable works. I'm not sure what all those replies are trying to get at. Of course there can continue to be an art-collecting market for human-made works made before the AI-art era.
For example, a woolen blanket made by my mother is much, much dearer to me than an equivalent woolen blanket made in a factory.
I see it very likely that certified-provenance human-art may command high prices in the future; but also for physical part, it's pretty trivial and doesn't have to be so posh and fancy - anything from street art to art galleries and art collectives/workshops.
The problem isn't for existing artists (except in terms of ethical issues)—it's for new/budding artists, who will have to contend with challenges to the authenticity of their work. How can they prove that they put their blood and sweat into making a piece of artwork by hand when an AI could've generated something equally passable?
The art is always human-made. Missing attribution doesn't mean "machine made it". Humans developed the AI, ran the AI, created the training data, picked the non-horrible sample from the output, etc.
If an artist who studied Van Gogh intensely made an original piece in Van Gogh's style, does it have equal artistic merit as one of Van Gogh's originals?
AI is just going to be another aspect of provenance, and if it's found out an art piece had AI used to make it, it won't be forgotten or ignored.
The same holds for AI art.
People will pay more for the original because it's rare.
I find it similar to watching chess. AI/Chess engines will always be better than humans, but it's exciting to watch a human play to see what they can do.
Whatever line you draw will leave several canonical figures on the wrong side of it. Maybe that's OK and you'll draw your line in the sand. Just don't pretend it's a non-controversial consensus.
I find "what is art" discussions impossible to resolve and untangle from prejudices and biases. The least problematic answer I've found comes from John Carey, paraphrasing: "art is whatever someone has decided it to be". In other words, if I decide something is art, then it is.
The problem shifts slightly to a more interesting way to pose this question: why some art has more value than some other art (or even "does some art has more value than some other art")? Equally difficult to resolve but more prone to highlight prejudices and biases mentioned above.
I'm curious about the possibility of the commissioner role for this situation.
From my point of view, the AI is the artist and prompting an AI to produce a specific image is akin to commissioning an artwork.
I think we have moved beyond human agency, and creation of art is reduced to the simpler constituents, the roles of artist and (if there is one) commissioner. The request to make an artpiece can also come from a machine.
Also, we're comparing an art 150 years in the making (with its schools, philosophies, heroes) with one in its infancy.
Oh, for sure. As for the rest, well I might've also missed the whole writing prompt thing. But I've always felt as if reverse engineering the thing.
As a historical note, this argument is over 50 years old. In 1965, an engineer at Bell Labs generated pseudo-random artworks inspired by famous artist Piet Mondrian. He found that a) people couldn't tell if the artwork was generated by Mondrian or the computer, and b) people generally liked the computer's art better.
https://www.historyofinformation.com/detail.php?entryid=4437
I feel like the majority of "art" we consume falls in the second category.
Would you say the same of general intelligence? IE, general intelligence (artificial or otherwise) requires an agent?
When I started doing this, I was uncomfortable thinking of myself as the author of the image in any way. I'm okay with it now. There's some skill I am expressing, even if not much, and I am connecting with an audience.
To take a selfie, human intervention is minimal: someone decides they want to take a selfie; take dozens of them, pretty much at random; choose one they like; publish it.
Similarities with the stable diffusion process are striking. And yet, could you consider a selfie "artistic"? Are all selfies artistic? What if Annie Leibovitz or Marina Abramovic take a selfie?
What if Marina Abramovic generates a million images of a rabbit with stable diffusion and papers a whole warehouse with them? Is this art?
Edit: syntax.
https://www.npr.org/sections/thetwo-way/2017/09/12/550417823...
But most selfies probably aren't high value art.
If Marina Abramovic were to take a bunch of selfies, or generate a bunch of AI images, and exhibit them, that might be considered high value art. A famous artist can literally tape a banana to a wall and it'd be considered high art:
https://www.vogue.com/article/the-120000-art-basel-banana-ex...
But does that make it more artistic than something generated by AI which was coaxed out by a human applying their expertise?
In my pitch I proposed generous royalties and cash compensation to the artists. All of them essentially said that they weren't taking commissions. I'm sure this is in part due to the fact that I have never been involved in the graphic novel world, so they'd be taking a chance on someone new. Still, it seemed there was no amount I could pay to get someone to work with me.
Now I am reviving this graphic novel idea and still looking for someone I can pay to help me work on the project. However, the more synthography advances as a technology the more it seems like I should explore it as a path for this project. It would help me develop the book much faster than I can on my own. At minimum it could help me get the story board in place.
Maybe some day synthography can help artists scale up their output, in the same way Michelangelo developed a workshop of apprentices. I would be more than happy to pay an artist's AI apprentice if I can't work directly with the human.
There is a way we can pit some corporations against the others to help with this though: Train models exclusively on content by megacorps like Disney, then claim it's fair use.
Those guys lobbied to get copyright extended for ages for own profit; for once they could help protect the ordinary artist.
Training models from scratch only works with massive (labeled) datasets covering a massive data distribution. With language models, the datasets being used are quickly approaching "all known written text" sizes. Training a model from scratch on Microsoft's internal code, with not only its precious intellectual property, but also its technical debt. Code at Microsoft is not going to get close to covering the broad range of styles that a coder could possibly use. The model will possibly diverge without enough data, as it needs to see a given "example usage" in multiple different contexts before it can learn it.
My current understanding is that deep NN's are quite good at modeling an underlying distribution of data without needing any priors hard-coded about that dataset. But! They need to see a whole lot more of it than an adult human would. Several orders of magnitude more. - and they need to see accurate labels about 75-80% of the time.
He actually did disclose this [1], at least to the art fair administration:
> Olga Robak, a spokeswoman for the Colorado Department of Agriculture, which oversees the state fair, said Mr. Allen had adequately disclosed Midjourney’s involvement when submitting his piece; the category’s rules allow any “artistic practice that uses digital technology as part of the creative or presentation process.” The two category judges did not know that Midjourney was an A.I. program, she said, but both subsequently told her that they would have awarded Mr. Allen the top prize even if they had.
[1] https://www.nytimes.com/2022/09/02/technology/ai-artificial-...
Nit: 2 weeks ago there was an article about getting SD running on an iPhone. The app isn't very time- or battery-efficient, but it just barely works on my (previous-gen) device: https://news.ycombinator.com/item?id=33539192
Still, 2 minutes on my phone. I remember Bryce taking half an hour to render lower resolution images back in the late 90s.
That said, you can just rent a GPU instead of buying it right away. It's not very expensive, and can be cheaper overall depending on your usage. It's also the only sane way to experiment with training as you need a lot of attempts.
This stuff can really devour any hardware you can possibly throw at it.
I agree, this is an underrated option and I'm not sure why it doesn't get more attention. I've been using stable diffusion on a cloud machine, and, say $20 will get you very far - probably thousands of images generated and a few sessions of fine tuning. It's way more affordable that its apparent popularity would suggest
This development seems parallel to me: AI does not seem to be inventing, even in the limited sense of closely customizing for a purpose. It is remixing the well established. If machines can help with remixing and reapplying the old, that's all to the good. Humans still own the new.
Real-time heads up display
Keeps the danger away
But only for the things that already ruined your dayIt's like the author (and others here) see the artist and art as separate entities, as if art is just regular labor that gets separated from a worker
What makes Kafka good? Not just the words on the page, or the prose behind the words on the page, but the person who wrote it and his own unique life experiences. Same for art
The fact that some people see art as just another product of labor and not something more personal is incredibly chilling and solipsistic
Not only this, but long term if we start to separate the creator from the creation, we no longer will control the terms of the creation. Imagine a hypothetical world where everyone forgets how to actually make art, and the AI art generators are rigged to try not to show any regime-critical art
Will AI replace art? I doubt it, but the way we're talking about it is chilling
I'm writing this at 5am in a weird anxiety-induced sleep-procrastination HN session so I'm having trouble pinpointing and articulating the je ne sais quoi, but maybe some examples will meander gradually towards it: I've read many novels that I thought were good and knew nothing of the authors but the name (isn't this part of the point of a nom de plume?) and I've read novels I thought were middling, sometimes by the very same names. I know nothing about the artists of street murals I occasionally walk by, and I know nothing about the artists of non-mural graffiti tags along walls and overpasses on the very same roads, and I know nothing about the artists of phallic scrawls on the walls of the bathroom stalls I use along the way, but there's a pretty clear ordering of their artistic merits based on some ethereal non-quantitative aesthetic quality despite all being ostensibly "art". I'm not an art buff, but when I first encountered the works of Caravaggio, I found them to be a master class in form and light that resonated with me on some level, and I still thought that after eventually reading about his brawls and exile due to the murder/manslaughter of Tommasoni. Sometimes it'll take more than a year between discovering a musician or band that I enjoy, listening through their oeuvre, and finally reading their Wikipedia article (or Soundcloud/Twitter bio, for the sufficiently obscure). For some, it never happens. In the hoi polloi art form of standup comedy, I can still laugh hard at Louis C.K. or John Mulaney or Bill Cosby despite knowing stories about their lives off of the stage that paint them in a bad light, and I can laugh just as hard at the tight 5 or crowdwork or even the recorded album of someone I've never heard of, if the material and delivery have that non-quantitative quality of, well, being funny.
I don't find my art consumption habits or perspective troubling or chilling. I'm willing to celebrate what humanity is capable of without celebrating a specific subset of humans, or even knowing what subset of humans are worth celebrating.
It's like how I try hard to write clean yet robust code at work, but am not bothered that even when I hit the mark, it will only be appreciated by those within my corporate walls, and fade when I or the organization am gone. You can run the "git blame", sure, but "it's just code" - today's "big wins" are tomorrow's "legacy garbage". Ashes to ashes, dust to dust. But hey, I've struggled with lifelong chronic solipsism, so maybe this perspective is chilling to you nevertheless. I'm a real human, there's a "me" in here. I don't like to think that what makes me smile - and why/how - is any more a canary for our march towards oblivion than anyone else's smile-inducers.
As for me, I just feel that it should be looked at the same way that photography is looked at; that it's a new medium all to itself. You wouldn't enter a photograph in an oil painting competition, so you shouldn't enter AI art into a drawing/sketch art competition either. It would also be useful if there was an acknowledgement that AI was used in some elements of the art, but like as not there will always need to be a human agent that makes corrections in what is created.
If these advancements are a result of systematic ignoring and abuse of image copyrights, is it ethical to continue to use them?
That said, as much as it is 'image copyright abuse' problem, it's still directly connected to AI art, so denouncing all AI art is justified. I don't see these issues being resolved any time soon. This massive scale exploitation will go on, as it continues to exist in legal gray area and not get enough legal attention and pushback (when a system abuses billions of works, and individual artists might be small and disparate, it may be hard for pushback to accumulate, or even happen at all), and then it might just get a green light to continue functioning the way it does. Sure, some 'more ethical' systems might pop up, but that won't change the bigger picture.
I think what makes me uneasy about these AI art generators is that they can trivially assimilate a new style like this and commoditize it. That's why so many of the most successful prompts are "X in the style of ${artist}".
Similarily, computer+artist will always produce more compelling work than computers alone.
That has not been true for several years now.
https://en.wikipedia.org/wiki/Advanced_chess#In_future_studi...
Has this changed in the past 5 years?
I'm having trouble finding actual tournament results.
This was true maybe a month ago, but DreamBooth makes this possible
I've read people making comics and generate many different scenes say v4 now allows consistent character design.
I took ten high-quality pictures off Pixiv, spent five minutes cropping them, and ran Dreambooth overnight on a 3090 — this turned out to be enormous overkill, the final model was massively overtrained. Intermediate ones were fine.
The output is perfect; even higher quality than what the input model was already good at. This remains true even if I wander well outside the scope of the input pictures, e.g. different art styles, or Shamiko in a spacesuit wandering Mars, or...
https://usercontent.irccloud-cdn.com/file/w300/l9rBbcFo
Trust me when I say that 'magical girl shamiko' is not usually a thing. The technology works.
https://cdn.discordapp.com/attachments/807260261747785759/10...
How about Shamiko on a highly improbable forest run? A completely different art style / "medium", in a completely different pose from any of the input images. Mind you, it's lower quality; that's mostly because I didn't do any cleanup or img2img work.
It's still not as good as the real thing. She's missing her tail in both of these pictures — though she had it in others — but even if I filtered for that, it's missing the expressivity it usually has. Shamiko's an emotional girl, and a lot of that shows in tail behaviour...
I doubt I need to tell you that there is no way to say "Jealous Shamiko, with tail curled protectively around Momo". Not yet, at any rate.
Why can't this be done? It also seems like an object-composition problem, so assuming the model has some concept of "Momo" like it does for "forest" it would seem to be possible? Is this just a limitation of the Dreambooth finetuning process?
Also if you happen to have more samples of the Dreambooth output, could you share? (I want to see that Shamiko in a spacesuit wandering mars...)
I'm interested to know how well diffusion models can generalize across style (I guess this could be tested by keeping a fixed prompt and varying the initial seed state). Have diffusion models successfully learned some latent space for "style"? This doesn't matter for real-world objects since all apples look pretty much the same, but it matters a lot for 2D art where artists usually have a unique style (which I guess would be defined by proportions, palette, etc.).
The tail is also highly stylised, frequently taking up completely inorganic poses such as “jagged pikachu style shock/surprise line”.
One or the other of those might be manageable, at least by generating fifty pictures and picking the best. Both in combination means human input is absolutely necessary. Which doesn’t make the AI useless, by any means; it’s fully capable of acting as a collaborator.
So this seems to be more of a incremental improvement, from a layman's view. I'm sure it's technologically impressive, but it looks like it won't be generating ref sheets. Thank you for testing!
First there was photography, then there was digital art, then there was nfts, now there is ai art - after each of these more and more people appreciate the inherent creativity of the non-deterministic nature of the human spirit.
Enjoy your cheap Beksinski imitations if you will, but it's better to go see his work in person in Poland. You will understand the meaning of art.
Genuine question: what would change your mind on this? Or is "human-made" part of your definition of the word "art", and as such there is nothing that would change your mind?
Additionally, do you have at least a vague idea where the line is between "human-made" and "machine-made", in the spectrum that spans "oil painting" - "photography" - "photoshop" - ... - "dall-e"? Maybe you're considering a completely different spectrum - I'd be very interested to hear your thoughts on that.
Please note, I'm not trying to argue "it's impossible to have a hard threshold so everything goes", I'm just trying to better understand where the art world is on these topics, as I'm mostly in the tech bubble otherwise.
Most of us (art world) can look at an image (a.i. or otherwise) and immediately recall the artist(s) it was mostly influenced by. As a common example, this is why most dark surreal a.i. output feels like a botched Beksinski piece, however the tragic life of Beksinski informed his technical stylistic depth. Dall-E is a purely (practically speaking as compared to other forms) deterministic system. Even photography requires the human eye to compose, light, design, filter, granulate, texturize, etc.
Edit: Not sure if I exactly answered your question, but I think I've tried to explain how there's a direct connection between the human experience and novelty in style, A.i. undercuts it, but a conscious system cultivating new styles based on its experience would be the closest thing to change my mind.
I agree with the problem, on platforms like AWS you'll even need to send a manual request so they would let you use instances which can run SD. On the other hand there's already something like replicate.com, which allows you to run the SD like an API. I hope there will be more services like this.
For me AI art is the same thing. It can give you serviceable material in a very fast time, but that's it. It's a Big Mac of the art industry. It's not gonna kill artists like how McDonald's didn't kill restaurants or even rival fast food chains or gourmet hamburger shops, because the only thing redeemable in AI art is you get variations of well-known styles (putting "trending on ArtStation" in the prompt to make it look good is a thing. let that sink in).
How is art defined? How is it defined by artists and how is it defined by consumers. Some artists may want to define it such that it's made by people. Much fewer consumers would make that a requirement (aside from some art collectors). Consumers will care much less of how some art was generated and they will dictate demand.
Artists should rejoice, in my opinion, mastering this type of tool could dramatically increase their productivity, and thus their salaries.
I take it you haven't been keeping up. Stable Diffusion 2 and Midjourney are getting really, really impressive. Probably a 10x increase in quality over the original DALL-E already.
Take a look at Jodorowsky's Tron (Tron in the style of Jodorowsky's Dune): https://www.djfood.org/fantasy-jodorowsky-tron-visualisation...
At first yes, faces and hands weren’t perfect but in the recent models and especially with model blending I can now generate super realistic photos of humans with perfect faces and eyes and perfect hands. Basically what is possible and the quality is rocketing forward right now and I can’t even begin to imagine where it will be in a year.
I would also add, the advances in prompt engineering, in-painting and image to image generation mean getting exactly the result you want with the composition you want is also very possible. If you’ve only generated a few images then you really have no idea of all of the tools and options you have. Over the last few months I’ve probably made 5000-10000 images and my own personal skill level for being able to get what I want has gone up massively.
I work in the video games "industry", we need concepts and mockups, but a much larger part of the production is assets, meshes, textures, materials, shaders...
We're seeing huge improvements, from photogrammetry to simplified production pipelines and more powerful light models.
ML based tools will also help to iterate faster, but the key point is control, human control.
Everyone who's dismissing this because it imitates the style of another artist (as if that was a hard technical limitation rather than a specific prompt) or lacks the je nais se quoi of "real" art is missing the point. No one would ever hire a concept artist capable of doing this when they can hire someone for a fraction of the salary to write prompts that return acceptable results in a fraction of the time.
Pretty much like power tools for carpenters. The house won't build itself, but the process will be improved.
Do they? Show me something truly novel either in terms of art or message created by Hollywood in the last decade. Art and entertainment are risk averse industries - novelty is risk, genre is safe, and within genre you have well defined aesthetics, themes and material from which to derive new iterations on a theme.
The mashups in the blog posted upthread (particularly the Giger/Henson posts[0]) look far more visually striking to me than most modern sci-fi. I want to see that movie based on the look alone, and the look is literally just "what if Dark Crystal, but xenomorphs?"
Salaries do not increase with productivity. Rather, they tend to remain constant or even decrease as productivity rises, as technology makes that level of productivity easier to hire for, and thus less costly for the employer.
How many office jobs have been replaced by computers?
More productive. And I would argue still mostly underpaid given the value they create.
> How many office jobs have been replaced by computers?
Most of them?
However - I do wonder if this will increase the value of non-digital arts and crafts. They can't easily be substituted for - which is probably one of the key factors in establishing value in the art market.
Barriers to entry? Training and hard work are barriers to entry? So the next time when you go to the dentist, tell him that you can fix your tooth with a text prompt.
Yes, assuming there is no substitute.
I have wondered if merely being behind the curve will be the Achilles heel of AI art -- nobody wants a bunch of old-looking crap, and if you're on the cutting edge, there isn't a large enough body of work to train the AI from, so it can never catch up; i.e., 2030 art might not look like 2020 art, and you may never have a current-enough labelled dataset to look "current"
I do hope AI art will make game and film production cheaper, but at the same time I suspect for commercial ads and a large part of graphic design, AI won't be that useful, since it is "imprecise" (can't fit the art in exactly such and such a position in a beautiful way -- or even, think of a dashed line with misaligned corner-dashes), or ads will look old and "wrong".
Just thinking out loud.
So they're taking about "animation"?
This and image-to-image generation (more in a moment)
A reader will see there's more writing, you don't need the parenthetical.
I found this funny as well, but on reflection I think the author is talking about the specific effect of changing the seed but not the prompt. To me it highlights a weakness of AI imagery — no intent, no progression, no story.
Eh, yet. I can see where the prompt just gets extended to, "in a hilarious series events in the style of buster keaton" and all of a sudden your vampire-toothed anime furry in cyberpunk clothes is flying down the side of a building, being saved at the very end by a sunshade they flop into. But I don't know if I'm going to live long enough where AI makes new Buster Keaton shorts using commonplace elements around them in new and exciting ways. Like would AI know that you can hang from the hands of a clock, which move over the course of a day? to me that's some creativity that was years ahead of its time.
Now that I've attempted to predict the far-from-now future, I'm sure we'll see this in 6 months.
If Copilot is any indication, we're really not far from some tool which can take a spec as a prompt and output an entire software system - front end, back end, DB layout, configuration - which sort of fulfills that spec.
I understand that photography generates waste images. I just don’t feel like it’s the same.
In my view, in the past it was believed that to be an artist and produce the output you admire, you had to learn things. You had to be patient, and spend thousands of hours learning a skill. But I don't think it starts and ends with just 10,000 hours of brushstrokes. All that time and effort likely causes other neurological changes owing to the situations that being a traditional artist forced you into, like being exposed to other people and communicating with an audience.
The ability to generate art with prompts bypasses all that time and effort. Using Stable Diffusion, I have made incredible pictures (that I don't consider art), but at the end of the day I still feel like just another person chasing yet another rush of dopamine, the same as it's always been. It's an enabling device. I have never been able to summon the will to practice drawing or painting for enough time to see results, and that hasn't changed. The only difference is that I have the end results, which is mostly what I wanted in the first place, and it's pretty much all I'll ever get with this setup.
But it is often repeated that "focusing on the end results" when trying to learn a creative skill results in bad outcomes. It is a character flaw that someone has to overcome to get better. Now I feel that 70% of the road to "the end result" can be bypassed without the need to work towards that additional knowledge and virtue.
When artists speak of AI damaging the integrity of their profession, I think this is part of what they mean. For the first time in human history, it is now possible to be a better "artist" without growing into a better person, or developing a wider-scale notion of what existence or the world outside is. The folk assumption that growing up as an artist changes you in fundamental ways (by forcing you to study the craft and observe the world for what it is) has been violated.
The people who treat art as a profession or hobby to work at over thousands of hours, for example, will probably share a lot of significant interactions and bondmaking as their skill gradually improves over the years; not so with a person that can generate pseudo-masterpieces within minutes.
I'm going to guess that the long-term downsides of AI art are going to be baked into cultural norms instead of the limits of silicon. There has already been significant backlash from art circles when tech-oriented people outside of the scene train models in the style of people like Kim Jung Gi, who had just recently passed away. Many people's rules about "art" have gone unwritten for so long because this current situation was unthinkable only a decade ago. Regardless of intent, being seen as plagiarizing the style of someone who is no longer alive isn't a good look.
Still, what I always want to keep buried in my subconscious about how AI art is used and treated is that my insatiable interest in the subject is fueled by my lack of (ability/drive/bravery) to express myself through my own actions, whereas my embarrassment is spared by delegating to a machine to do the work for me and guess what the inner workings of my imagination look like. If this is something I need to overcome, it's only going to get harder for me to do so as the state of the field advances further.
The word that keeps coming to mind when I look over the things I generated is "intent". It doesn't feel like I "made" the things I did, but that I "found" them. There's not a lot of insight to be gleaned except the keywords that each subset of popular culture finds the most useful or interesting in regards to prompt engineering during a particular timeframe. My preoccupation when generating pictures is to find a result that looks nice, not something that carries intentional meaning or context that's greater than the sum of its parts.
This is a big part in why I enjoy vastly different kinds of music creation. When I work with a traditional DAW and build a track from scratch, I'm sculpting and crafting an artwork. When I work with modular synthesizers and generative compositional tools (AI or not) I feel like I'm navigating sonic landscapes to find interesting destinations.
AI art is the latter, and while it's fun, it's a different feeling during creation all together. In music, I can typically hear which the artist was doing in a given song. With Stable Diffusion and DallE, I don't know that I can as reliably. This is why the divide is so interesting to me. We have to radically different avenues, each with their own merits in how fulfilling they can be, to choose from which produce similar results.
I will admit, there is something to the reward of making something hand crafted that at least to the artist, adds value, even if that value isn't (typically) transferable to the end user.
Like, this isn’t some kind of hack’s approach to art. Literally every working lyricist that I know uses a rhyming dictionary from time to time. Why? It’s fucking work! There are deadlines!
With regards to AI visual art… the best results I’ve seen have been from people with strong visual art skills. Aesthetic choices still matter a lot when you’re tweaking the prompt, model, hyperparameters, etc.
Engineers figured out how to scrape sites, decode and filter tags, classify images, and feed them into the training dataset.
All the original artist did was publish an image online to be seen by other people.
It'd be like Google compensating people for putting things online that the search engine links to.
are there occasionally people who say this is something which may upend art? of course. but is there a “panic”?
the people this type of clickbait immediately grabs are those who seem to enjoy explaining why “the panicky and the hysterical” are wrong wrong wrong.
but ultimately most of the time they seem to be attacking a ghost.
Digital artists doing commissioned work in the style of some other more famous artist? Bespoke stock photographers?
It seems these jobs will simply use the new tools, and sell a more manual version when needed.
Similarly, AI-remixed¹ music is getting better and better. I want to know how the author will feel when their music² is used to build a model that allows anyone to create infinite music in their style and beyond, much of which will be enjoyed far more widely (because it will be free or royalty-free) than the source material ever will.
¹ It doesn't seem fair to call it "generated" or "created" when it'd be useless without an enormous corpus of source material powering the thing.
Right now, your music requires attribution and can't be used for commercial purposes.
Should visual artists have the same ability to choose the rights they reserve?
Sure, but fair use still exists. If someone takes just a single drum hit, reverses it, and coats it in distortion and reverb, but some how I find out they did it, I can sue them. But I can sue over anything, that doesn't mean I'd win. I'm pretty confident they would, and would agree they should because it's fair use. I think training an AI model is fair use as well.
> Should visual artists have the same ability to choose the rights they reserve?
Yes, and within the bounds of fair use as musicians that I've stated above. If I photoshop out a single eye, let's call it 50x100 pixels of a 5000x5000 image, and use it to make a neat mash up art, I think that is transformative enough to be fair use. Similarly, I think the AI model is transformative enough.
In either case, there are limits. If someone makes a model fine tuned on my specific music, that's wrong. Similarly, if someone makes a model fine tuned on an artists specific work without their permission, that's wrong. Using an artist name (music or drawing) in a prompt is gray, but I personally think leans on being acceptable, especially if that artist is dead.
The reality is that most art creators are getting paid by marketing departments to create essentially stock art for various marketing initiatives.
it's likely not particularly enjoyable work, but is probably necessary for living for a lot of unknown artists. If that market is categorically removed from the equation, these artists will have serious issues going forward.
It's like the old argument of workers getting redundant by automation. Sure, some people will find new jobs in other deciplines, but not everyone.
Though to be fair, current models aren't there yet. They might get there eventually, and that's the fear if a lot of people.
> In the brain, information is transmitted through the release of chemical neurotransmitters at the junction, or synapse, between neurons. This process is stochastic in nature, which is believed to be a key feature in how the brain operates.
1. the medium & technique: what technique is being used to put the artwork into this world? I beg you all, please go to a museum full of paintings (let's not forget that there are sculptures, installations, and so on) and take your time to look at them. you will find out that the digital representation of "the girl with the pearl earrings" is not even half the artwork. it's an oil painting and that has an incredible depth to it. see how the figure is put into the frame - who cares about a hypothetical room she is in - the important part is what is being shown and what is being kept from us. the light and the dark parts.. I could go on... please try to expose yourself to art in person and find out how much more than colour values a painting is! (brush strokes, depth of colour, size, varnish, etc)
2. originality: why have someone paint in the style of someone else? people who are not into art often don't understand that a style is always bound to a person and the time they surface. art after 1945 had many ways of reflecting and processing the horrors of the nazi regime, the holocaust, etc the invention of photography changed the art, but not because craftsman went out of business, but because the subject of the painting changed. before photography abstract art would not have been possible (although artists already changed their styles to less realistic) hyper realyism on the other hand evolved long after photography was invented....
3. history and provenance of the artwork and the artist:
I'm getting tired typing this all out on my phone :) I think i have to redo this in a more concise way with my computer or maybe pen and paper...