Synthesizing Obama: Learning Lip Sync from Audio
grail.cs.washington.edu
grail.cs.washington.edu
He explained his system at a fairly detailed level for a short 5 minute video. He was breaking down the source clip into very small snippets of time, and with some algorithm, giving each snippet a single value, or a score. Then, as his realtime audio was analyzed in the same manner, the source snippet with the closest score to the realtime snippet was output, both audio and video. I've always wondered what happened to that project, and I've since forgotten the creator or the name of his project.
* deepfake can superimpose a face on another face
* Adobe is developing a program to make a voice say anything (from a couple of hours of voice input)
* this can sync you mouth to the new voice generated by adobe.
Combine all three.
People aren't dumb and it'll only take a few good fakes to get people to distrust videos not attached to a reputable source.
People are going to have their lives destroyed by intelligence services over this.
there's a paradoxical effect to it though. If these technologies become so pervasive that they're not distinguishable from real footage they'll lose their entire value, especially in a legal case. If you could arbitrarily create fake dna tests there'd be no point to pay attention to the results.
https://www.timesofisrael.com/when-a-job-interview-turns-int...
> In Israel, too, the results of a polygraph test are inadmissible in criminal court but in civil court they are fair game if the person being tested agrees to it in advance.
For an illustration of the second point, try to recall your earliest memory. For me, the first time someone asked me that question I had an answer which was then fixed in my brain as "my first memory". If I try to recall the memory, it's very apparent that there's nothing real left. It's just a memory of a memory, and any specifics are imagined extrapolations of home video and photos which I can actually remember. When this happens for more recent memories, the process is much more transparent.
Check out [1] for a discussion of current 'allowable' priming practices. Courts and legal scholars aren't blind to these tricks.
[1] https://ilr.law.uiowa.edu/print/volume-101-issue-2/the-inves...
Use an look-alike actor to make your own video. Then faceswap with deepfake. Sound can be taken from Adobe Voco.
But you need a thousands of pictures and hours of audio for this to work. And it's yet probably stil recognizable. But in a few year, this can be quite scary.