Nvidia’s Implicit Warping is a potentially powerful deepfake technique
metaphysic.ai
metaphysic.ai
Not sure I will care either, because by that time I will have been locked in my living room for weeks playing my customized VR role playing game where Han Solo and I wreak havoc in the Far Cry 5 universe with my Dovakin powers while being chased by a all-female bounty hunter crew commanded by a 1982 Phoebe Cates.
You can fool more people with a fake article and $1000 in Twitter click farm spend than you could ever dream to with a deepfake.
Like we've already seen entire elections undermined with basic internet trolling, the problem is already here, but somehow people are overlooking that and fixated on the most fanciful version of it?
One the one hand, you can't fault the masses for buying the goods they're sold, and on the other hand you can't fault the sellers for maximizing the apparent quality of their products, but somewhere inbetween all of that broad mindedness and beauty there has to be something at fault for creating a world where actual human beings are not good enough to participate as equals and the playing field is pay to win on a scale that boggles the imagination.
Strong disagree. Viral content is way more effective.
And also, imagine what you can do with $1000 in ad spend and a deepfake.
> And also, imagine what you can do with $1000 in ad spend and a deepfake.
Way less. A deepfake detracts from your mission. If you deepfake Joe Biden saying "I'm going to destroy this country as instructed by my masters" and you actually gain traction, suddenly you've thrust your subversion into the spotlight and start reaching people adversarial to your goal.
News sources are going to start digging to find where the clip came from, the White House is going to respond, the video is going to start getting analyzed to death.
-
If instead you register a dime a dozen domain like "america4evernews.com" and write a crackpot article about how Joe Biden is actually working for our enemies and you have it on good authority that he's being controlled by puppet masters who want him to destroy America, you'll find an army of people who want to believe.
They don't need sources, they'll avoid sharing it with the "sheeple who believe MSM" and stick to their echo chambers. It's a strictly better outcome.
People don't seem to understand that modern misinformation is not about fooling everyone, it's about fooling the right people. You're goal isn't to reach every set of eyeballs, it's to reach the eyeballs that are easily fooled so that they come in conflict with those who are not, thus entrenching both sides against each other.
-
In some ways it's like a virus: if it's too strong it draws attention to itself early and can't spread easily. If instead it causes subtle symptoms that compound, it can spread very widely before it's even noticed, and use that infected base to expand even further.
If you bring evidence you're introducing a place for a counter attack to apply leverage.
If instead you make completely baseless claims on obvious false pretenses, you've actually made things more difficult to counter because only trivial counterproofs exist, which have to be dismissed to believe the false claim in the first place.
-
Take COVID vaccine deaths for example. Imagine I baselessly say that the medical field is lying and 50% of COVID deaths are actually vaccine complications.
For someone to believe that, they must completely distrust any official numbers on COVID deaths... so once they've fallen for the lie, how do you convince them otherwise? The only counterproofs are the trivial to find sources of data that they already had to dismiss to believe me in the first place. Suddenly I've implanted a self-enforcing lie that entrenches its believers against anyone who isn't in their echo chamber.
The root of all this is straightfoward enough: there is nothing stronger than requiring someone to disbelieve their own beliefs to counter your disinformation. If you add a deepfake, you've added something outside of their belief system to attack, so you're weakening the attempt. People simply do not like to be wrong about things they think they've figured out.
> Imagine I [...] For someone to believe that, they must completely distrust any [data]
Do you think this is this like a 419 scam where saying something a bit outrageous sorts out the gullible and bypasses the wary or do you think that your claim can somehow hijack a credulous person long enough so that they make that mental recategorization of the data sources and are stuck?
The people falling for the obvious nonsense are self filtering just like people falling for obvious 419 scams.
But as you grow the base that believes in your disinformation, you gain real people who are fully convinced of these things, and the effect of that is a force multiplier.
People talk, and if people's self-held beliefs are the strongest reinforcement, the second strongest is those we surround ourselves with. If someone falls for this stuff and starts talking to the spouse, now someone close to them is pushing this agenda. People can start nudging their friends to be more skeptical.
It's not going to be a 1:1 conversion: a lot of people close to them will push back, but remember, this is all based on absolutely no proof, so it can twist itself to fit any box. People can moderate the story to avoid pushback: "Oh you know I'm not an anti-vaxxer... but I did heard that vaccine has a lot of complications", and maybe they connect that to a real article about a myocarditis case, and now maybe they're not pushing my original lie of "50% of deaths", but I've planted an suggestion in a rather moderate person using a chain of gullible people.
And something especially effective about this is the fact that, while the most brazen aspects of disinformation hit less intelligent people hardest (https://news.ku.edu/2020/04/28/study-shows-vulnerable-popula...)
Once you start to make inroads with increasingly better educated groups via the network effect, they tend to not want to believe they're wrong. Highly intelligent people can be more susceptible to some aspects of disinformation in this way: https://www.theguardian.com/books/2019/apr/01/why-smart-peop...
That lends itself to increasingly authoritative figures becoming deeply entrenched in those campaigns, leading to things like... https://wapp.capitol.tn.gov/apps/BillInfo/Default.aspx?BillN...
-
Overall I've said this before, everyone is dreaming of AI dystopias rooted in things like deepfakes putting us in a post-truth era, or AI gaining sentience and deciding it doesn't need humans...
The reality is so much more boring, yet already in progress. We're starting to embed blackbox ML models trained on biased or flawed data into the root of society.
ML already dictates what a large number of people are exposed to via social media. ML is starting to work its way into crime fighting. We gate access to services behind ML models that are allowed to just deny us access. How long before ML is allowed to start messing with credit ratings?
And yet getting models to "explain" their reasoning is a field of study that's completely behind all of these things. You can remove race from a dataset and ML will still gladly start codifying race into its decisions via proxies like zipcodes, after all it has no concept of morality or equality: it's just a giant shredder for data.
Right now a glorified bag of linear regressions is posing much more of an effective danger than T1000s ever will. But since that's not as captivating instead we see a ton of gnashing of teeth about the ethics of general intelligence, or how we need to regulate the ability to make fake videos, rather than boring things like "let's restrict ML from as many institutional frameworks as possible"
It’s not only not captivating, it’s downright inconvenient. If I’m at a TED talk I don’t want to hear about how ML models (some of which my company has deployed) are causing real world harms __right now__ through automation and black box discrimination. If you read Nick Bostrom’s Superintelligence it spends laughably little time pondering the fact that AI will likely lead to a world of serfs and Trillionaires.
No, people want to hear about how we might get Terminator/Skynet in 30 years if we’re not careful. Note that these problems are already complicated by ill-defined concepts like sentience, consciousness and intelligence, the definitions of which suck all of the oxygen out of the room before practical real-world harms can be discussed.
And yes, I do not think adding a deep fake of some big shot saying or doing X would work better than just repeating the lie that they say or do X.
if you ask me, this 'tradition' is the essence of imperialism (or just one amongs other techniques necessary to have an empire)
Oh, absolutely. And we're not talking decades here. This tech will likely be ready to rock within a few election cycles
The term is just spot on.
Facemask and soundmask are good terms for a democratized warping of visual and audio data.
Considering how facemasking and soundmasking have multiple contextual meanings, I think deepfake is the perfect word. If you need more precision, then you can use the phrase "visual deepfake" or "audio deepfake".
That being said for smaller countries with less internet access (and less education) this has already become a big problem (lots of cases in elections in Africa) so I think once again it comes down to education: we must inoculate people to bullshit like this by giving them the tools to spot it and future things like it.
Call me a cynic, but I believe humans, myself included, are easy to fool. If the face looks good enough, plenty of people will fail to spot the Invisible Gorilla.
"Facial Masking" is not bad either, quite accurate, but I'm not sure it provides as broad a description than "Deepfake Puppetry".
Either way, it is important to get a term that is both accurate and resonating with the audience to ensure that the general public understands the serious potential damage for this technology (as with any powerful tool, can be used for good or evil).
It not only captures the face but also the body, and maybe the voice. Facial masking is more programmer friendly.
window.removeEventListener('wheel', getEventListeners(window)['wheel'][0].listener); metaphysic.ai##+js(aeld, wheel)
Done.Why do people think it's a good idea to circumvent the scrolling behavior that the developers of the browser probably have spent hundred of hours perfecting over the years?
Or uBlockOrigin in "medium mode" or higher incorporates JS blocking on a per-site basis.
And it required extra work to yield this negative behavior!
So how in the world does this end up happening? How could a team be so profoundly detached from reality? This site's behavior is so incredibly ill considered that I marvel that people worked on that, people approved it, people said publish, etc, without one person stepping in and asking what in the world they were doing.
I switched to chrome, and that is just horrible!
I am curious why I am not seeing this on Safari (and apparently others on Firefox). I don't see any errors on Safari implying that some javascript is failing to load.
I would have expected the worse behavior on Safari... not Chrome.
Or am I just being nieve about frontend development? I really only do it for any personal projects so don't really have any insight into it professionally
One side perk of not using Chrome is that there's a correlation between only testing on Chrome and producing code that we're better off not interacting with.
I just find it interesting because I look at the apple product pages. While a bit janky they work as intended. But they are also applying that effect very deliberately and not to an article for some reason. So maybe that is the distinction.
Like, if I had the time and resources to work in AI, stuff like substituting faces and even stable diffusion would be about the last things I would ever work on.
What would I start with? Something more like the MS Office software of the 1980s, only automated. I would have real-world applications that, wait for it, perform work so I don't have to. AKA automation.
TBH this stuff exhausts me to such a degree that I almost can't even follow it anymore. It's like living in a bizarro reality where nothing works anymore. A waking nightmare. A hellscape. Am I the only one who feels this way?
But that's what Stable Diffusion et al is: automation. It's just not automation for what you spend your time on, but it is automation for what a countless amount of people spend their time on, doing stock photos/drawings/clip art for endless amount of articles and other generic videos.
I also think that the current "creative automation" comes from a perspective where many people have said for a long time that computers will of course be able to do boring, repeatable jobs like counting numbers and what not, but it will never be able to do the job of an "artist". But now, it seems like at least a subset of "artists" will be out of job unless they find a way to work with the new tools rather than against them.
Artists however wont have to worry about ai. Ai music already exists but people still support mainly real artists on streaming platforms and go to concerts, because provenance matters for artists and it doesn't for designers. Could an ai make a warhol? Probably not because what made warhols art popular was that he essentially worked outside of the training set and provided something previously unseen. Machine learning is bound to the training set. You can make generic corporate bathroom art for hotels or filling empty picture frames with it, but there will still be real artists and galleries and concerts and museums, because often times people value the provenance of the artist much more than even the work itself.
Just like how "Calculator" used to be a job title for humans who manually performed calculations.
I don't think that's what's going on. The top pop singers are generally already singing things written by other people against accompaniment written by other people. I think by far the biggest reason few people are listening to AI music is it's just not as good as human music yet?
I don't listen to my favorite music because of who made it but because I like it.
If music that I like started bring made by AI then it's listen to it without hesitation.
In fact a lot of music bring made today is already a collaboration between humans and machines and had been so for a long time.
A lot of world is grinding in pain due to extremely bad software and the money and brain power keeps pouring everywhere but there. Well not entirely.. there was a lot of money thrown on these bad applications, but it evaporated due to software services companies subpar engineering.
I had the fortune/misfortune of moving furniture for 3 years right out of college 20 years ago to support my internet business at the time. I saw how the vast majority of people toil away their lives to make rent and child support each month. That experience shattered my will to such a degree that I came out of it a different person.
My concern is that the divide between working poor and techie riche is now so vast that they can't even see one another. If the wealthy and powerful could see, they would invest less in profitable schemes and more in shared prosperity. But they can't. So wealth inequality continues to grow unabated, with AI being just another tool to profit from another's labor or eat their lunch outright.
Plus, these are the things we hear about because they look flashy. There is plenty of work behind the scenes on applying these innovations to more 'practical' matters like automation.
Then I project when the answer is just under my nose.
It's good to be reminded from time to time that what we're asking for may already be here, we just need to realign our perception to see it. I truly believe that's what meditation/prayer and manifestation (magical thinking) are all about.
... but, I think you have really missed the point! Maybe you think government is here to help rather than to govern minds too! And that what is shown in the news is a good faith attempt to relay reality!
If you are managing the world, companies, etc - perception is everything! If you are able to control what people perceive, and they receive everything via a screen, well, who cares about truth? The imagery, the ideas - that needs to convince... and that is pretty much all that you need to manage the masses.
I can also say from experience that extreme negative feelings like "other people are doing things I don't find interesting and it makes me exhausted and miserable" were, for me, a sign of clinical depression.
The catch is that I feel most depression is environmental today. It's singularly exhausting to struggle while watching people who have the means to enact real innovation and change squander their potential on yet another gimmick.
The only thing that's really helped me was to realize that it's all a gimmick. Life itself is a divine comedy. We each determine our own definition of meaning since science can't provide one.
Using the book/movie Contact as an example, basically my philosophy has shifted away from Ellie Arroway (Jodie Foster) and more towards Palmer Joss (Matthew McConaughey). In it, he wrote a book called "Losing Faith: The Search for Meaning in the Age of Reason":
https://www.youtube.com/watch?v=HFcHpamkHII
This spiritual battle we're engaged in between science and faith has been with us since the beginning. It's crushing down on us harder and harder now as science works to stamp out the last vestiges of our individuality and humanity. That sentiment might not make a lot of sense to many on this site, but it's the daily lived experience of billions of people forced to spend nearly the entirety of their lives toiling under subjugation (working towards another person's life purpose) to survive.
You're also right about interest and motivation. I find AI to be perhaps the last frontier, since there's a chance it could explain consciousness and maybe even give us access to something like an all-knowing oracle or even a means to contact aliens. It's pretty much the most interesting thing there is, and why I got into computer programming in the first place. I'm just sad quite literally that I squandered so much time on running the rat race and never got a chance to contribute and "get real work done".
It's becoming increasingly clear that many of these techniques that for the past 10 years that have been derided "pointless" or novelty now have real applications.
For one example, the automotive industry - ignore the hype on autonomy, computer vision is already delivering real benefits for active safety systems. I use github copilot every day - its not perfect but good enough to add value to my workflow. Apple's automated tagging of my photo library via computer vision allows me to discover hundreds of images of my life and family I'd forgotten all about. Stable diffusion can clearly replace an artist in some cases, ignoring the moral/ethical issues.
I'm extremely excited for future of all this, frankly. The first step into such a new paradigm is always hard - people made the exact same "home computers are pointless" arguments in the late 70s/early 80s. I don't think anyone agrees with that anymore...
Wow I thought this copilot thing was just like a joke. People use it to write software? I'm really curious, now. Can you share examples of the code you're writing with it?
If it gets it wrong sometimes I don't really care, all the time it takes is pressing tab to complete or not to complete if suggestion is garbage. It's just an extension to the tab key's functionality, which is why it's so easy to use every single day when it integrates to most text tools.
I also think if this is what we can have today, the future of code completion tools is very exciting.
Could you please share an example? I saw a description of copilot but couldn't imagine what it might be useful for.
Neural game upscaling is a huge win too. The neural-TAAU upscalers (DLSS 2.x and XeSS) perform a lot better than FSR still, even compared to FSR 2.0/2.1. And there will be other things they can figure out how to ML-accelerate as adoption continues, I'm sure. It enables some solutions to hard problems that don't have good deterministic algorithms, and you can run a lot bigger models than you can without acceleration.
also fabrice bellard (of course) wrote a neural compressor... https://bellard.org/nncp/
it's also likely going to be useful in level design and asset creation as well, although of course as we're seeing now with Stable Diffusion there's some interesting legal questions with content creation.
optical flow engine (not neural) is another big win for computer vision stuff, too, offloads a few tough tasks to hardware, like object tracking and motion estimation. hooking into that with zoneminder or something would be awesome.
I'm also stoked for shader execution reordering too. This is similar to what Intel calls "ray binning", it basically is a best-effort "re-alignment" of the thread (state,etc) to the most similar execution group for memory read alignment/coalescing. I think one or another raytracing implementation (Intel or NVIDIA) they were coalescing based on material.
Intel talks about theirs here: https://www.youtube.com/watch?v=SA1yvWs3lHU
Ada whitepaper: https://images.nvidia.com/aem-dam/Solutions/geforce/ada/nvid...
I am interested to hear what AMD is doing with RDNA3, supposedly there is a new RDNA3 ISA instruction for matrix acceleration but rumors are it's not a full high-performance matrix unit like CDNA and >=Turing? I don't know why you'd add an instruction without some hardware acceleration though. So maybe less than the other implementations... which is a little disappointing.
Perfecting these two things opens whole universes of potential.
Ask a robot to “do the dishes”. It has to know what that means, find the dishes, find the sink, the on/off mechanism for the water. These are all language and vision tasks.
The balance, navigation, picking/placing etc seems like a minor subroutine fed by vision metadata.
There's significant potential value in networks that can regenerate their input after processing.
It can be used to detect confusion - if a network can compress + decompress a piece of input, any significant differences can be detected and you can tell that there's something that the network does not understand.
This sort of validation might be useful, if you don't want your self driving car to confidently attempt to pass under a trailer that it hasn't noticed - which tends to kill the driver.
I saddens me to see here these sort of posts.
If you meant "trust some things if they fit some kind of criteria" you mean something different than "can't trust anything."
I grew up with the prevailing idea of "never trust what you see online", yet today many have deep and misplaced trust in online media, influencers, and personalities.
I think a little more distrust is needed
Not quite. Deep fakes will undermine people's trust in anything digital they see or hear. It bodes ill for technology not for people per-se.
In other words, those who should be most terrified by the implications of deep-fake technologies and AI are those most heavily invested in digital technology. I sense that in some ironic way, suddenly the "boot is on the other foot".
This is inherent of our species. The idea that you should distrust anything you see or hear is an alien concept and simply not pragmatic.
"Misinformation has always existed and always will exist"
False equivalence. The online situation is brand new. Any citizen able to spread massive amounts of fake news to lots of people on the cheap is a brand new capability.
"the solution is education"
No, it isn't. Every study shows that highly educated people are also gullible enough to be manipulated with fake news. Which makes total sense, as absolutely nobody has the time to fact-check and do a deep background check on the 5 zillion pieces of information they see on a given day.
We should actually be much more proactive about technology: it is not an unalloyed good, and an unwillingness to consider its downsides leads to naive designs. Even leaving bad actors aside (which I don't recommend doing), what works well in a group of 10,000 users may not scale to one billion users. Where one scale has no effect on social cohesion, another scale may have a tremendously deleterious effect.
We've been doing these experiments in search and social for years now. Taking lessons from that and applying it to the next great wave of AI innovations seems like a Good Idea to me.
Looking back, what technology would you have retroactively stopped? Tetraethyl lead? Perhaps. Nuclear? Cable TV? ANNs? The Internet? Drones?
Also, the financial incentives aren't aligned. Certainly it makes sense to hold back technology to avoid embarrassment if you're a trillion dollar company; you have more to lose than gain. However, if you're a scrappy startup, it makes way more sense to roll the dice.
This really isn't true. We regularly do cost-benefit analyses for business; we don't skip them because they're not perfect predictors. We do market analysis and all kinds of customer deep dives to perfect UI/UX and customer response. We've created targeted dopamine-delivery services that are continually refined to maximize impact.
All of this implies a certain kind of ability to evaluate. Looking at tradeoffs and potential uses of technology is well within our abilities, and we should do it. We should be more skeptical of human nature, look harder at the extremes and edge cases, and work towards mitigating the risks. Will it be perfect? No. Will it be helpful? Yes.
It's very valuable to discuss the potential dangers of new technologies and how we might mitigate them.
> AI might put false ideas in peoples minds
Like speech or writing or media?
> no no not the ideas, its the medium, the delivery method. you see the fakes will trick people into thinking the lies are real.
oh that's it? just deception + scale.
Who are you quoting in your arguments here? I did not make those arguments and I cannot find someone else in this thread who made those arguments. Perhaps you are creating a strawman to argue against?
Re: Fire: I wasn’t around for it, but it’s safe to assume people discovered fire was dangerous around the same time they discovered fire, wayyyy before cities existed and before humans decided to harness it.
Re: Social harms from AI: you should ask yourself about the harms that can come from automated decision making. We’re automating decisions for policing, for hiring, for delivering posts on social media, for delivering political advertisements on social media, etc. We’re using AI research for profiling, for improving bomb drones, for getting children to spend time and money on games and social media, etc. I’m sure you agree at least one of these are harmful.
Re: Social harm from ‘deepfakes’: It’s currently costly to create a convincing fake image, audio, or video. It’s easy to extrapolate that it will be easy to make convincing fakes in the near future. It’s easy to see that can cause harm, especially since people are already tricked by obvious photoshops and deep-fakes.
In the US at least, there are political attack ads rampant today that use altered media.
I find it difficult to find a generous interpretation of someone who thinks we should proactively shut-down any discussion about potential and current harms from AI.
Yet the toothpick is a little bit less dangerous than the atomic bomb.
We shouldn't blindly accept every new "tech" as "progress", the fact that we can build it doesn't mean it'll be beneficial for us
Because it means that finally, after decades of us being blinded by the wow-factor of technology, hackers and engineers - we who are responsible for creating the next wave of technology - are starting to ask grown-up questions about whether some things are such good ideas.
That healthy scepticism doesn't have to mean pessimism, or Luddite rejection of "progress". That's enormous progress of a different kind - a shift from purely technical progress to a better balance of spiritual and social progress.
Since then, "the victors write the history books", including formative influences on children's thinking.
It's encouraging if we manage to collectively scrape our way back from that, and have genuine concerns -- despite the noise of fashionable posturing, "influencer" conflicts of interest, etc.
> obliterated by the frenzied influx during the dotcom gold rush.
Something changed in mid 1990s, something that was more than just the commercialisation and Eternal September. Perhaps it was millenial angst but a certain darkness and nihilism crept in as if William Gibson stopped trying to imagine future dystopias because the future had "caught up" and we were living in it. I think that's when things actually stopped progressing. After that, everything that had been written as a warning became a blueprint.
> Since then, "the victors write the history books", including formative influences on children's thinking.
80s and 90s media moguls only owned the channels, and there was cursory regulation. When you own all the platforms, and devices that people use, and shape the content they see from school-age onwards it is no longer "media" but total mind control.
> It's encouraging if we manage to collectively scrape our way back from that
It may not be "back", but it may be somewhere better than here. Each generation finds it's own voice. I have some optimism in the kids since 2000 - it's like they're born knowing "The cake is a lie". They're just not quite sure yet what to do about it.
Here's how it might look from the client side (computed style at Inspect->Style):
.elementor-875 .elementor-element.elementor-element-230321ec { text-align: left; color: black; font-family: roboto; font-size: 21px; font-weight: 400; line-height: 1.6em; }