This is better than nothing, but it's not going to provide immunity against AI fakes trending and having impact before they're identified as AI.
We don't need the metaphysical solution to the problem of detecting AI videos for the rest of time. Certainly, it's fairly easy to make something that mostly works most of the time. Enough to be very, very useful.
The parent post's worry is warranted, IMO.
The point is that, on average, individual experiences are improving, decade by decade. Humanity is doing a remarkable job.
Drugs are out of control. Homeless are everywhere. No one has interests in anything. No one is having kids. All jobs are going to be gone soon. Colleges can't teach (it's all AI cheating now). People are Gang Robbing stores. Cartels are killing hundreds daily. Fraud is out of control. We have 2 maybe 3 world wars going on simultaneously now. Prices are skyrocketing.
Yeah I get why you say "pretty much". lol PS good luck buying a house
My daughter's English professor is now requiring people to hand write their essays during class. So at least there is that.
Yet, in the grand scheme of things, the world has gotten better in that time. Which is how you know these people are wrong, at least most of the time.
The challenge with positive news is that you have to go seek it out. It will rarely come to you. And then you have to greet it without cynicism, which many are incapable of.
"Works most of the time" isn't good enough here.
Things are not perfectly fine how things are now. AI slop is destroying the internet. Tons of grifters are earning tons of money off YouTube by brainwashing millions of people with AI slop, including my mom. YouTube needs to do something and this seems feasible and far better than doing nothing.
I also think the false positive rate is going to be far lower than you think - especially if YouTube sets a caution threshold.
I'm open to other solutions but if you propose we just keep what we have now, then you are proposing an absolute disaster.
It's not like human-generated content is made of carbon and AI-generated content is made of silicon and the science of chemistry can unambiguously tell them apart. If you asked a million humans and a million LLMs to write a sentence on a specific subject, it's not implausible that one of the LLMs and one of the humans would output the exact same sentence. Maybe more than one.
A thing that can take only the output and accurately tell you if it was AI-generated or not is therefore impossible, because if it said no it would be wrong when the LLM generates that sentence, but if it said yes it would be wrong when a human generates the exact same sentence.
All it can do is try to calculate a probability. But then what do you want to do with that? Suppose the probability it estimates for some content is 45%, and that probability estimate is an accurate measure of the true probability, i.e. can't be improved when the only information you have is the content itself. Do you want to ban the 55% of that content which is human-generated, or allow the 45% which is AI-generated?
I get the idea: get 10k each samples of human data and AI data, train a simple classifier until it gets 99.9999% accuracy or <10k false negatives per day at your scale, ship it as a screening tool.
Is such tool feasible at all with current state of AI technology, or is it just a reasonable take from the past that may not be so reasonable anymore?
"If a creator doesn’t specify whether or not they used AI, but our systems detect significant photorealistic AI use, we will now automatically apply a label."
The issue is, that's not a thing. AI-generated content and human-generated content have significant overlap. No amount of training data can allow you to distinguish them with that level of accuracy because many outputs exist that could have been generated by either one. Additional training data allows you to say that the probability is 55.0374% plus or minus 0.0001, rather than only being able to say that it's 55% plus or minus 5%. It can tell you with greater precision exactly how ambiguous it is. What it can't do is remove the ambiguity.
Email spam filtering can clearly cause reputational harm too.
Only if it actually works
It could work "well enough" for YouTube to consider it a success while still harming a fairly large number of content creators.
I'm sure many content creators' videos will be labelled as AI generated. For good reason.
Problem is that at YouTube's scale the remaining "some of the time" ends up being a collossal figure. On top of that, YouTube's effective monopoly position magnifies the damage done by false positives.
It's not just from AI either. Video creation used to require a fancy camera and a above average internet connection. Now the whole world has that so we're seeing a lot of low quality profit seeking content on any platform where there is money to be made. There was a GitHub repo with 100s of low quality PRs because people thought it would boost their job prospects.
As for false positive, the most straightforward path seems to be to let stuff slide unless you are really sure. Maybe that slightly rewards players like Kling because they keep the invisible watermarks for their own use, and that of the CCP,but not third parties. NBD.
It's not like catching everything is that important. YouTube isn't claiming this is perfect. And I don't know that anyone need this to be perfect. It's not like even the best photorealistic video creation tools don't have plenty of tells anyway.
This doesn't seem like ZeroGPT at all. Having a flag or not having flag on a YouTube short is low stakes. Its not like it's being sold as a solution for something high stakes like academic grading.
Cryptographically verifiable provenance and chain of custody is going to be necessary to get to the human only stuff, before long, but the good AI stuff will be better. Just a matter of time, at this point.
Unfortunately that could still be true while labeling all human-crafted content as AI-created.