Lucasfilm hires YouTuber who specializes in deepfaking big-budget movies
theverge.com
theverge.com
Obviously, the labor-intensive nature of today's CGI techniques drive up production costs. Meanwhile, the deepfakes on YouTube provide a convincing enough rendition of likenesses without actual actors, all produced on consumer-level GPUs. This presents a huge potential to save costs and the benefits are clearly enticing to film productions.
As Hollywood gravitates towards blockbuster franchises, productions will want to bring the same ensemble of actors (or at least their likenesses) as long as possible. While moviegoers may be unsettled by seeing "reanimated" dead actors like in Rouge One, they still may hope to see franchise actors to look consistently youthful or attractive on screen. Deepfakes may be more relied upon to provide that effect.
https://www.youtube.com/watch?v=dHSTWepkp_M
Just look at it, the CGI DeNiro looks like from the Polar Express.
In movies you don't usually have the luxury of choosing the original face you want to replace, but for people who make memes (or porn), you commonly choose a source video featuring someone that resembles the person you want to put in.
Hiring an actor that roughly resembles the person you're trying to deep fake seems doable.
It's an intuitive result even ignoring the specific model. It takes less information to go to and from similar faces than it does two completely dissimilar ones.
Deepfake definitely has a place in this space, in that it can do things with 1/100th the effort in a scalable way, but it has a lot more to catch up on with traditional modeling than what some folks appear to be thinking. rendering with goodies likes SSS, AO, etc. gives magical results which are hard to achieve any other way. And as soon as you get a little bit complicated in what you're trying to create, at least the currently existing neural network models fall apart are are just not applicable. Take this video for example, which was very manually modeled: https://www.youtube.com/watch?v=BC2dRkm8ATU Deepfake is a long, long way to taking a stab at things like this.
I agree with your complaints about DeepFake but imo it did a far better job de-aging DeNiro. To me, the release version had a lot of "old man" cues that the DeepFake one didn't, such as the jowls and heavy wrinkles.
Not to say the deepfake isn't seriously impressive. But I very much prefer the netflix version.
If Deniro appears as Scorsese wanted him to appear then the deep fake (that I do find subjectively better) is actually aiming at the wrong target. It seems like a flawed comparison.
It works when you are trying to do that on actor that are no longer with us, but what about actors that are still alive?
But maybe it doesn't apply to parody(?)
It reminds me of the ‘what’s in the box’ 2009 short made with available cgi assets https://m.youtube.com/watch?v=IU_reTt7Hj4
This is going to be one more dimension to misinformation on the net.
Now you will see videos of politicians saying something and even then you cannot be sure whether this is actual video or a fake.
Or if you see a quote claiming to be from Emmanuel Macron or Boris Johnson, you can see it was released with a digital signature from a Guardian journalist & they add whatever date/time/location details they want to validate the information.
If instead I release something (on Twitter or wherever) that I say is a screen-capture of my TV, _I_ add the metadata (originally seen on BBC on 27 Jul, 1:55pm) and sign and then you know that you only trust it as much as you trust _my_ reputation rather than the BBC.
Wouldn't solve the problem entirely but it might create a bit of an audit trail for stuff and encourage people not to trust unvetted material.
In that scenario, you'd want people to say "but wait, there's no signature on this, it could be fake" and then only trust the video as much as you trust the source (not the claimed source).
Any visual artifact can be mocked, so we end up with the same problem as clickbait titles, where the conclusion one arrives at from just a title can be disproved, but it doesn't prevent the false information from going viral.
What good is it to say "that video you saw was fake!" after the video has spread around and done the damage already?
It's hard to come up with a solution to this problem just because the solution has to preempt the problem. A cryptographic visual artifact _could_ work, but it's still likely that misinformation via deep-fakes will cause problems for society at large.
Websites like Twitter adopting a "blue tick" for a validated profile on their platform though is a model people seem to get. If we had some equivalent of a "blue tick" at a user-agent level, e.g. a for your browser to take a signature and display it in a standardised, human way to say "this video is signed by bbc.co.uk" it could work. (With a similar model for user-agents elsewhere e.g. you'd probably need adoption in apps like WhatsApp to get traction.)
The other side of it (like privacy discussions) is how much the average person will care — tabloid journalism often skirt the borders of what they can get away with at the moment & they nominally have a duty currently to only write factual information. If Fox News or the Daily Mail release videos and put their own signature to it, then you arguably lend them legitimacy ("it's on the news so it must be true. It's signed by them and all!").
That's the mood in the algorithm hole twitter put me in to.
So, make what you will of that.
But why store the signature in a blockchain? If you do not trust the certificates in the first place, the storage location won't make any difference. And if you trust the certificates, the storage location is completely irrelevant. Because the certificates alone provide the trust.
For the same reason as certificate transparency logs; you want to avoid trusting something that has a history of certifying false statements. You also need to handle throwaways, so it's definitely not sufficient (and might turn out to not be necessary once a complete solution is found), but it does seem useful.
Disclaimer: I am an armchair crypto fan. Not an authority.
Similar to how anonymous 'tip-off' stories with protected sources work in general at the moment — the media outlet put their own reputation on the line on the basis of the source & we trust (to a certain degree) reputable news outlets to validate & vet their sources correctly. This is true for stuff that's easily forgable at the moment, e.g. a whistle-blower releasing documents.
NFT for news!!
This changes when individual people with no resources at all can make convincing fakes and wield it as a weapon to sow disinformation, to have it then picked up by "major" media, all information on the net becomes pretty useless.
Sometimes a random video pops up and people believe it shows X and then it propagates but it's something completely different (e.g. people beating up immigrants on the streets in northern Italy -> traditional krampus celebrations; Junker drunk at some event -> Junker suffers from lombalgia; Berlusconi mimicing a sexual act on some woman -> it was a comedian skit; Britney Spears sex tape -> it's a random pornstar ...)
People will learn to be doubtful of random internet videos just as they have learned to be doubtful of random internet articles.
Or not, since they haven't yet, but it's not a qualitative change.
It's easy to lie to people so long as you're saying things that validate their shitty emotions. Conversely, it's extremely hard to tell people the truth when it goes against their shitty emotions.
Emotions can't be shitty.
You can miscalibrate your emotional responses to situations, much as you can sear your conscience.
But the emotions themselves - the full range are valid human feelings, from fury to transcendent joy.
Those feelings are valid, but the effect they have on other people is not. Dealing with valid emotions in a way that doesn't harm other people can be incredibly difficult, especially when those emotions put harm front and center.
Our emotions are valid, but our behaviors are not. It behooves us to mind our emotional states when they cause problems for other people. Often, that will coincidentally bring about an emotional state we prefer as well, but often with unpleasant transitions.
I didn't say anything about behaviors, because it was emotions that were labeled as shitty.
I didn't bother to go into the distinction between emotions and the bad behavior they often give rise to, so thanks for explaining that.
Emotions as such are good, from sorrow to rage to joy.
Each one is a good response to some situations, as far as I can tell.
Your emotional response can be misplaced, so that you experience an inappropriate emotional in some situations.
As noted elsewhere in the thread, you can also be inspired by your emotions to inappropriate (and just terrible) behavior.
The emotions themselves, though, are not shitty. Recognizing them and understanding where they're coming from can be tremendously helpful in aligning your actions with reality and your own values, and even in discovering what your own values are.
I share this perspective not out of a sense of superiority but in the hope that it helps someone else avoid my mistakes.
E.g. saying "I think people of type X are bad so I feel hatred towards them" is a good emotion but misplaced, vs is a bad emotion? I don't see the useful distinction.
It's called "repression" and it can really mess you up.
I don't mean that emotional responses are automatically good - I mean that emotions, as abstract entities, are inherently good.
I'm not so sure I'd call "hate" an emotion, so I'll try a different example.
If I'm angry at someone because they told me I wrote a buffer overflow (and I actually did), that's an unreasonable and unhealthy response. The fact that I feel anger over it should impel me to introspection and working out why I'm angry over useful technical feedback. From there, I can move on to personal change so that I'm no longer inclined to be angry about good, helpful input.
If I'm angry at someone because they sexually assaulted my wife, that's a healthy anger response. They've mistreated her horribly and my anger on her behalf should push me to protect her and seek justice.
How exactly I act on that anger matters, though - physically intervening so they cannot continue to hurt my wife then calling the police would be good.
Pulling out a pistol and shooting them would not.
Does that clarify my thoughts at all?
Calling all emotions axiomatically good, but also saying that they are "unhealthy" if they're "inappropriate responses" still seems unnecessary.
created by 12 people. spread by many thousands.
I think the opposite is true. This isn’t a historical perspective, you’re using logic to speculate. Historically speaking, there have been fakes that reached huge numbers of people, and they were more damaging then than they are today because they were more believable; the public had not yet conceived that photos could be faked, and it was not possible to see evidence of fakery. Today, everyone knows photos and videos can be faked.
I don’t know of any deep fake videos yet that have tricked a large number of people or been used for political purposes. Maybe it has happened, I don’t know, do you know? But there have been lots of influential faked photos. Just Google a little to find hundreds of historical examples of famous and misleading doctored photos. (Lots of overlap in these lists, because some of the photos are famous).
https://www.cc.gatech.edu/~beki/cs4001/history.pdf
https://delmarwatsonphotos.com/photographs/famous-photograph...
https://www.quora.com/What-historical-photos-are-highly-misl...
https://www.businessinsider.com/fake-photos-history-2011-8
https://www.ranker.com/list/historic-images-that-were-retouc...
https://www.ba-bamail.com/content.aspx?emailid=29607
https://www.pinterest.com/yosomono/faked-images-everyone-thi...
[1] https://en.wikipedia.org/wiki/Hippolyte_Bayard#Self_Portrait...
I don't buy these sky is falling arguments.
Deep fake video will have about as much impact as photoshop has had.
Back to text and the good'ol credibility of the messenger. Digital commodified journalists and now the need for credibility will let them get out of anonymity again.
A video needs to actually be watched, that takes more effort, and then it would be widely debunked as fake. The "fake news" memes are usually at least partially true which helps convince people that the misinformation is legit.
As far as you know. By definition, wouldn’t its successful use mean you didn’t know it was successfully used?
But totally agreed it needs to be common knowledge that everything digital can potentially be bogus. This stuff should be taught in schools from an early age honestly.
Neural Radiance Fields (NeRF): https://www.youtube.com/watch?v=JuH79E8rdKc&t=5s
Good, clean, and reliable information is expensive and needs a fair bit of work. I can see why most people have forgotten that but it might come back them and then this will be way less of a problem.
This is actually already the case with Biden and Trump video clips even without being deepfaked. Often they are presented out of context to the point of completely reversing reality. It's helpful to assume any clip is fake by default, especially if it's a viral one that makes one side look bad.
By the time deepfakes are common, it'll be best practice to assume fake by default.
Nothing new under the sun
People will be less likely to believe leaks or supposed hot mic recordings, but the majority of what politicians and other public figures say happens in public view which makes it difficult to fake.
You already don't really know if a video you see on Facebook has been carefully re-cut to change the meaning or tone of what the speaker was saying, so I think the fact that we can more easily wholesale create videos doesn't really change much. If you want to know if something is real the best option is still to cross reference multiple sources and if possible multiple recordings.
Yep. You don't need CGI if you can search a large population for someone who looks like insert-public-figure-here and make up the difference with makeup.
The problem is when everything you have ever learned is suspect.
When has this not been the case?
It's about to become pretty low-effort for a random person with an axe to grind to do that.
It's not that you can no longer believe anything published online - it's that video evidence without provenance was relatively reasonable to trust for a few years, and it's about to stop being so.
as Ricky Gervais said in his infamous Golden Globes speech, "This show should just be me coming out going "Well Done, Netflix, you win everything. Well done."
When I was working in indie game development, I wondered if you could use deepfakes as a voice actor. Basically get someone famous/good voice with infinite voice lines, without having to pay for studio time. Obviously, you would need them to sign-off on using their voice for commercial purposes.
[1]: https://www.gamesradar.com/witcher-3-mod-uses-ai-to-create-n...
"With barely contained terror. You drive a hard bargain."
And one can even imagine a program with emotional slider bars that lets a person listen to how a line sounds with different levels of inflection and then automatically inserts the appropriate markup for the settings the user selects.
The new version is almost ready to launch.
I've also got voice to voice conversion working, and I'm trying to make it real time. It's pretty close.
> James Dean, an iconic movie star who died in 1955 at the age of 24, has been cast in a new Vietnam-era action film called Finding Jack.
That's a good thing. More actors can now work.
https://youtu.be/A8TmqvTVQFQ?t=52
That said, one way it *does* work is with novel scenes filmed with a good impersonator, the outcome can be pretty remarkable:
https://www.tiktok.com/@deeptomcruise/video/6957456115315657...
https://www.youtube.com/watch?v=krAU3C9jhj8
As long as the creator wasn't making money from it.
It's one of the reasons why the franchise survived for so many years without new movies.
I don't understand the complain. It's a movie: what you saw never was believable.
It's trickery by design. Everything you see as been spliced together from multiple takes, purposefully framed, lighted and colorized. It's all fake but in a way so culturally ingrained that people don't even notice the deceit anymore. You think this continuous action you are watching?
although you are right about video 'evidence', editing, cuts, and carefully muted dialog can alter things to the point of being the opposite of what was being filmed - an unprovoked attack can become self-defense or vice-versa, etc. Again, deep-fakery is just another tool in that unsavory toolbox, not anything paradigm shifting.
Plenty of videos are candid or shot and released by a 3rd party.
This could work for press releases and such, but not for videos with headlines like "CEO CAUGHT KICKING A BABY IN THE FACE" or "UNDERCOVER INVESTIGATION: PRESIDENT ADMITS ALIENS ARE REAL".
Image formats could be updated with an 'original' layer so that if the news site wanted to crop or edit the content, the original would still be avaliable for comparison.
Given all of the hardware design going into high speed hashing on asic, shouldn't be that hard to find a component to do the work.
I know there are cameras that digitally sign the data. However, you hardly ever get to see raw footage. It is always edited.
Deepfacelab has some tutorials to explain what is actually done: https://github.com/iperov/DeepFaceLab
I understand this is a mistake, we all make them, but where are the editors at The Verge? If a reader can find this with a cursory look-over, shouldn't they find it also? I couldn't have handed this in as a high school essay, so I'd imagine it wouldn't get past an editor at a large magazine. Maybe it's just me but at least from what I see personally there's loads of errors small and larger in the news these days. It's weird.
Proofreading is difficult because your mind subsconsciously fixes the text you are processing. Editors need to focus hard, but as we all know it's tough to maintain focus while doing monotonous tasks.
It's not a big problem, the meaning is still there, but it seems like a trend to me with online journalism at least, and I was wondering if others felt similarly or if I'm just being unfair.
I as an amateur have proofread texts for my friends many times and I often missed very visible mistakes. But it's not my job of course, maybe specialists have some methods beyond reading carefully.
From what little I know about journalism (very little), editors at traditional media had a very opinionated stance about language and punctuation. It is to the degree that they're not merely finding errors like these, they're also suggesting rewrites for clarity, etc.
I, too, notice many simple errors like these that make me think that editor must not be as valuable a role as it was in the past. As the cost of communication dropped toward zero, an editor role becomes a more significant cost, maybe. Would the cumulative effect of errors like this one be enough to impact the readership of The Verge?
There’s no ground truth to compare with other than the Luke skywalker character in the original Star Wars movies and which one looks most realistic
That said, the facial animations for Leia in another video he made (they did a digital version of Leia for Rogue One, he deepfaked on top of that) actually improved with the deepfake version.
Do we have different actors deepfaked in for different markets?
Do actors even act anymore? Do companies just pay actors for their likeness and do the rest?
Do we get to a point where you can choose who is acting in a film you're watching?
How are deepfaked voices? Can we substitute audio as well?
I understand that deepfakes are still a bit of a manual process, but presumably that will change.
Part of it is inertia and random celebrity status, having an attractive or interesting face etc., but part of it is also the raw knowledge of when to apply certain microexpressions, how to gesture etc. i.e. how to do the acting itself. To be a convincing, charismatic etc. actor it's not enough to wear a digital mask of a celebrity, the underlying actor still needs to act well. That may not be so important for certain types of shallow movies, but it certainly is for deeper drama films etc.
It's similar to today's text generation where you may be able to generate sports game reports, user's manuals or travel brochures etc. but not really those where you need high level decisions, like applying the appropriate expressions to a real-world event, taking into account all the context, like writing a poem about your feelings reflecting on some recent real-world event.
I'm not saying humans have a magical power that can't be implemented in silicon.
What I'm saying is that deepfakes as they are today are not sufficient to replace actors. You'd need a higher level puppeteering AI that would take the whole storyline and script into account to come up with the right ways to express the appropriate emotions at that moment in the film and could take the director's instructions regarding his vision of how the drama should unfold etc.
Theres a good documentary about this [1], that talks about replacing that actor, and when Fox comes on set and delivers the first line filmed ("You put a time machine... In a DeLorean?!"), it's hilarious, night and day difference.
[1] season 2, episode 1 of movies that made us www.netflix.com/us/title/80990849
I assume that would eventually get to generating a totally new person, rather than modeling after a specific actor who costs money.
Maybe it's just something someone has to explain to us what to look for. It could be a curse to know why some people think the deepfake is better. I once talked to a graphics artist about why they thought some effect looked bad because I didn't see it. They explained it in detail. Now I can't unsee what they were talking about and can easily spot that mistake.
I don't know. Maybe it's better not knowing and enjoying the results or knowing and each time you see it thinking: "Hey, they made that mistake"
Elsewhere in this thread, someone posted a similar comparison clip about The Irishman and the difference is even more pronounced there. De Niro’s eyes actually look the age they’re supposed to be in that scene. The original scene had two key problems, in my option; one, they decided not to touch the eye and focused on smoothening the skin (I assume this is because of the technical limitations of doing the deaging in painstaking CGI); two, and this is a little more of a mystery, is that they thought they could get away with using De Niro’s current 80 year old voice on a character that is supposed to be 35 (ages are ballpark numbers). The raspy voice of an old man is just not something you expect out of a supposedly much younger and healthier man. They should have just gotten a voice actor who could do a convincing impression of De Niro in the 70s and dubbbed him in.
The director's focus was on the most genuine De Niro expression, not on the most impressive young De Niro impression.
Scorsese doesn't make tech demos, he makes stories with characters, and he respects his actors a great deal.
What could've been done here is use another AI technology, deepfake for voice, basically, train it on young De Niro's voice, and then reproduce De Niro's own lines with his younger voice, using the original lines as input (not just text). No voice actor could match a performance THAT closely as AI can these days.
This was the big winner for me. The deepfake had "Tarkin" and "Leia" on screen. The originals had 3D GCI (very good CGI, but CGI) with dead eyes.
To further my anecdoe, I totally missed Leia's sequence the first watch through because I was watching the content and zoned out, forgetting to evaluate her face for deepfake pixels because it looked real enough for me to suspend disbelief.
I have never been able to look at Leia's face in the original Rogue One scene without forcing myself to say "they did their best, they'll redo it someday for the 8K release, until then grin and bear the dead eyes".
Here's another example from the same channel with a live actor (Robert Pattison as Batman):
https://www.youtube.com/watch?v=dmuYz0aZGgU
A careful examination will find all sorts of artifacts but an unprimed general populace wouldn't have any clue.
Truly amazing what can be done with consumer grade stuff nowadays.
Mind you, impersonating influential people in a convincing fashion is / has been a thing for a while now. I'm thinking of Forrest Gump hanging out with the president and the like.
(It's inspired by a book that I haven't read, so I'm not sure if the same idea we're discussing is also in the book... But it's definitely in the movie.)
I don't remember where I read that plot synopsis, so I might have some details wrong. (If it's a book I haven't read it, or it could have been someone's description of a story idea.)