Audio2Photoreal
github.com
github.com
I'm apprehensive about accepting nonverbal communication that a model has appended to a human source.
It's technologically impressive, but I'm failing to see the use. Can someone else enlighten me? I'm sure there's something I'm failing to see.
For example the npm installations of react show that old versions are still downloaded a lot in the last 7 days: https://www.npmjs.com/package/react?activeTab=versions
jQuery on a random CDN with stats also shows that many websites are likely never going to update it: https://www.jsdelivr.com/package/npm/jquery?tab=stats
You also sees that with system administrators: here is curl on debian from people who shares about it: https://qa.debian.org/popcon-graph.php?packages=curl&show_in...
End-state for Winamp vizualizers: synthesize an entire living world from the audio alone.
Thanks for letting me know.
It's not easy, it's a trained skill, though I guess one that's pretty easy to acquire accidentally by daydreaming and doing art. I think I actually trained myself out of it as part of "safety" training; I stopped relying on mental models of many things and started forming mental models of the checklists to get them into that state.
Yes, this could apply to them. No, it doesn't mean it always does.
This also doesn't really add to the point that it wouldn't be suited to human conversation. You rarely use NPCs to have human like chats, especially for several minutes. The few games that do would be Mass Effect or LA Noire, both of which use mocap to avoid the effect I'm referring to.
That being said, this is pretty awesome for a lot of use cases that are less interactive.
https://youtu.be/0pXBXIrB478?si=iQ5YtDPBSaq0ynsv
I honestly prefer the Titanic avatars though.
Whether you're trans or you just want to join a video call early in the morning without dressing up, the applications are endless.
In many situations we demand that people dress or present a certain way, just out of bullshit social expectations. This is one way to eat your cake and have it too.
This is for the metaverse.
I'm not sure if there would be any other potential use cases beyond these two. Or rather, I'm not able to think of them at so far.
it's simply not possible within the near future, even today zoom/teams video conferencing is somehow highly compressed and shit quality with just low res 2D video.
How feasible is it to imitate what this model and codebase is doing to use it in a commercial capacity?
Did they release the dataset?
It would also be nice if Facebook would consider making an API to give Heygen and Diarupt some competition, if they aren't going to allow commercial use.
Although there will probably be a bunch of people who become millionaires using this for their porn gf bot service who just don't care about license restrictions.
I wonder where it is headed.
Long term this type of work helps solve big problems even if the intermediate steps don’t produce exciting results.
As an example, early image generators were pretty uninteresting but today they are widely utilized and generally considered impressive. The thing that researchers in the field know that the public doesn’t is that there’s 100 boring steps before the exciting release, and some of the boring steps are very exciting on a technical level. Those intermediate achievements represent 99% of what machine learning research actually is and others in the field appreciate those works.