Valve's worried that AI-generated art is in a murky copyright state, and don't want to open themselves up to being sued.
Valve's worried that AI-generated art is in a murky copyright state, and don't want to open themselves up to being sued.
Three possiblilities:
1. It's just fictional. Probably written by a troll or generated by ChatGPT.
2. Steam refused to publish the game due to some obvious copyright issues (like they told Midjourney to generate superman or one-piece characters)
3. Steam is banning any AI generated assets.
My bet is 1 > 2 > 3.
The internet is flooded with content right now to the effect of "Valve might be doing this thing", but not one of those sources has actually reached out to Valve for comment. Instead they all cite a random commenter on Reddit (or they cite each other).
There are dozens of articles that are just summaries of this reddit thread with no further effort put into them, and that's pretty much the norm these days for a lot of content.
https://arstechnica.com/gaming/2023/06/steam-mods-reportedly...
I wanted to point out that journalists at least check their sources...
Might as well just replace journalists by AI at this point, to a large group of people (me included) they've all made themselves more hated and untrustworthy than a company, economist, politician or civil servant.
So it sounds like OP slapped some half-assed generated images into a game and tried to submit it. Valve now can't really trust someone that does that to have done any due diligence.
"I've been developing a game for a while now, and am near ready to release it on Steam. I'd prefer it not to be associated with my name (as in I'd prefer people googling my name and future employers being unable to find out i developed this game )."
Use copyrighted material in porn that you charge money for and you’ll get slapped with a lawsuit faster than you can load the home screen.
Ask Alex Jones...
from post:
> contains art assets generated by artificial intelligence that appears to be relying on copyrighted material owned by third parties.
So I'm guessing 2
"When a game comes up as problematic, it gets flagged a bunch by steam users, and there's a meeting that takes place where they decide if these things stay on steam"[1].
0: https://store.steampowered.com/app/1853200/TYRONE_vs_COPS/
Not making a verdict either way but I find it interesting. I'd like to know more about the internal discussion(s) that took place to establish their frameworks. Especially given the company is private.
Games with significant Crypto and AI art components bring significant risks to Valve in both legal and social contexts (99.9% of modern crypto-related projects are an intentional scam, and AI art is a legal minefield right now).
On the other hand, violence and pornography are much more accepted by society (in the context of fictional enterntainment).
Besides my already established biases towards AI: It's threatening to creative endeavors, not because it exists, but because it will impact the earning potential of creatives.
If so, the reaction I've seen is quite positive. Very unlikely though.
I'm sure they mostly just don't want to wind up in court with a lawyer being able to say that they let [blatant example here] get published on their store. So long as they can credibly claim that there was no way for them to tell something was in an objectionable category, I'd imagine they're fine with it.
Their rules, if you're curious: https://partner.steamgames.com/steamdirect
Sounds like due diligence.
Adhering to legal and copyright standards isn't a "cop-out"
I understand that OpenAI et al would like to assure all their investors and customers that there's nothing legally problematic with using an AI to launder away copyright infrigement, but we're going to need a few lawsuits to have the matter settled.
It is completely reasonable for Valve to forbid this until it is sorted out. Keep mind they are a company of IP creators, creating a marketplace for IP creators. The whole reason Steam was created was to establish a DRM that fought the piracy of Half Life. I am on the side of Valve in this.
Imagine taking a really dumb gig worker, showing him 10000 images, some of them with watermarks, and then telling him "draw a red car, kinda like the kind of images you saw". There's a decent chance you'll get a red car that looks nothing like any cars with the data set (original work), and yet he'll paint a memorable watermark on top because so many examples contained it, you said "kinda like the kind of images you saw", and he doesn't understand that the watermark isn't meant to be part of the picture. I believe that's whats happening.
It’s entirely possible for a diffusion model to produce an original work and yet still hallucinate a ‘shutterstock’ watermark onto it, in much the same way as GPT can hallucinate valid-looking citations for legal cases that never happened.
Producing (distorted) copies of images in the training data takes some real effort, and typically only occurs for images which are heavily repeated in the training data... Most of the complaints along these lines can be compared to complaints that cars cause massive bodily harm if you steer them into lightposts: The problem is easily preventable by not driving into a lightpost.
Generative AI cannot exist without pre-existing bodies of work created by human labor. It also displaces that labor and hurts the people whose content was a requirement for AI to exist. From this view, AI is not fair use.
If this is true, then ordinary copyright law means that AI-generated media cannot be used unless you have a release from every bit of training data you used. At least some of the currently existing AI:s were trained with datasets for which such releases are impossible, so they should not be used.
Also, for the love of god, do not use any of the AI coding assistants, or if you do, at least never publicly admit you do.
This should apply to humans as well then because brains ultimately do the exact same thing. Nobody creates art in a vaccuum.
Why do I say that? They want the developer to prove they only used material they created to do the training and Valve has the resources to follow that rule unlike the rest of us.
No, they only want the developer to "affirmatively confirm" it. It doesn't say anything about Valve demanding some sort of proof.
At first approximation, yeah, the risk of getting sued to bits might be roughly the same. But the upside is not.
They are obviously able to identify some copyright material.
Assuming this is even real, it may have more to do with preventing another 1989 video game crash resulting from the market being overwhelmed with crappy games.
Then again, most AAA games today are broken pieces of suck, so IDK.
They also care about their reputation amongst content-producers (game makers). Youtube faced this exact dynamic back in the day and have found it better to side with the large creators who care very much about protecting IP rights and so they exercise a heavy hand against copyrighted material.
And take the reputation hit that would go along with that. Valve's business is 30% technology and 70% the reputation of being much less untrustworthy than the alternatives. If they lose that they can close shop.
Valve might lose some reputation among the devs if they keep rejecting or removing games but that is far from reality, at least for now. Things perhaps might change if game studios starts picking up AI generated stuff more and more and Valve decides to blanket ban them. But even then the users of Steam is a juicy target for devs and it is a hard market to ignore.
But AI generated content is NOT banned. You just have to prove you have the copyright (or permission) for the training data.
The only reason it wouldn't be easy enough to provide is if you just scraped any available data set with a complete disregard for intellectual property.
With the AI we can at least be 100% certain of which input you trained it on and under which licens, making the whole a lot easier to deal with, as compared to humans. The liability is the same, but it's much easier to avoid legal implications, so why not play ball and ensure that you have the correct licenses?
It's not clear that you need any license to train on data in the vast majority of cases. Having a license to train on it won't guarantee that you can grant your users a license to any particular output, especially given the addition of user input. And most of the utility is in creating outputs that are indeed distinct.
So the answer to "why not play ball?" is: 1. It's not clear that it's legally required 2. It would be incredibly expensive and/or slow progress dramatically, or limit you to pre-existing licensed content (e.g. Adobe) which drastically reduces some types of capabilities 3. Given #1, for any company that doesn't have an Adobe-style library, "playing ball" is essentially betting the company that it will become legally required, because on top of developing an AI model you're going to have to become an expert content licensing and documentation studio
Notably: not all AI-generated content, but rather AI-generated content from models that were trained on material that's not owned by the person submitting the game.
Which seems like as much as you can hope for in a policy?
Given the whole "Stable Diffusion reproduces the Getty Images watermark" lawsuit[1] that's still ongoing, it's not an idle concern.
[1]: https://www.theverge.com/2023/1/17/23558516/ai-art-copyright...
Hard to see how any plausible outcome that would have that result for users of SD (if model training isn’t fair use, that’s definitely a blanket-liability issue for Stability.AI — and Midjourney, and OpenAI, and lots of people training their own models, either from scratch or fine-tuning, using others' copyright-protected works.
But “using a tool that violates copyright in the workflow” is not itself infringement; whether and in what situations prompting SD to produce output makes the output a violation of copyright (and whose) would be a completely different decision, and while Ibcan certainly see cases (such as deliberately seeking to reproduce a particulaflr copyright-protected element, like a character, from the source data) where it might be (irrespective of the copyright status of the model itself), I haven't seen anyone propose a rule that could be applied (much less an argument that would justify it as likely) based on copyright law that gets you to “used SD, in violation”.
Lots of blanket ethical arguments about using it, but that’s a different domain than law.
The speculative worst-case outcome for these tools that I've seen suggested is the legal system deciding that an image generated with them is a derivative work of every image that was used to train the model. Since none of these models were trained on images that they had rights releases for, this would mean they're incapable of outputting images that aren't infringing on the copyrights of vast numbers of people.
I can't say how likely that is to actually be the legal outcome, of course, but it seems like the sort of concern that might lead to Valve's policy here.
The Copyright Act does not expressly impose liability for contributory infringement. According to the U.S. Supreme Court, the "absence of such express language in the copyright statute does not preclude the imposition of liability for copyright infringements on certain parties who have not themselves engaged in the infringing activity.
One who knowingly induces, causes or materially contributes to copyright infringement, by another but who has not committed or participated in the infringing acts themselves, may be held liable as a contributory infringer if they had knowledge, or reason to know, of the infringement. See, e.g., Metro-Goldwyn-Mayer Studios Inc. v. Grokster, Ltd., 545 U.S. 913 (2005); Sony Corp. v. Universal City Studios, Inc., 464 U.S. 417 (1984).
IANAL. Considering Valve not only gives games a retail platform, has to approve games before sale, and takes a cut of that sale and assuming the reddit post isn't a lie then I am gonna guess Valve's probably well staffed legal dept decided not to take a seemingly iffy legal gamble on a game that probably wasn't going to rake in a ton of sales anyway.