Live data from Hacker News

Avatarify lets users run realtime deepfakes on live video calls

inputmag.com

81–90 of 127 posts

Re: Avatarify lets users run realtime deepfakes on live video calls

#82
post #81

I don’t understand why everything thinks they have to be on video for every single meeting. I never turn on my webcam. It makes no difference to the outcome of the call.

It is most important in calls with 6+ participants. It gives folks a chance to signal that they have something to say when many times a call may be dominated by one or two individuals.

Perhaps there are other approaches for that particular case such as a button to 'raise your hand', however, many people have friendly relationships with their colleagues and simply enjoy seeing them when conversing.

Re: Avatarify lets users run realtime deepfakes on live video calls

#83
post #38

Earlier quoted context omitted.

Real Time Voice Cloning certainly has iffy output, but it's probably the most popular because it provides the easiest plug-and-play experience with even a simple UI to get started. The author says he's working on a more polished toolkit called Resemble.AI, but I've never tried it. https://www.resemble.ai/ There's certainly a market out there for just beautifying existing repos to making it easier for non-scholars to…

See https://www.descript.com/lyrebird-ai for another one with an on-site demo.

[deleted]

Re: Avatarify lets users run realtime deepfakes on live video calls

#84

Hey guys! I'm one of the founders of Impressions, the first mobile deepfake app on a phone. Try it out and give us your thoughts. Right now it's out for IOS but Android is on route. Here's our website https://impressions.app

Very weak output I’ve seen so far. Trying another before passing judgment, but honestly this is so obviously fake output so far. Edit: now done three images, all well lit face, no dramatic movements, and the output is just terrible.

It really depends on lots of factors. Lighting, angle, face type, the celebrity you choose and your facial hair. If you're squinting your eyes for example, you won't get decent results. DM me your ID from under settings so I can give you credits to play with.

Re: Avatarify lets users run realtime deepfakes on live video calls

#85
post #81

I don’t understand why everything thinks they have to be on video for every single meeting. I never turn on my webcam. It makes no difference to the outcome of the call.

A lot of information can be transmitted through non-verbal communication (e.g. facial expressions). Do you look at the faces of other people on the call with you who do have their cameras turned on? If so, why?

Re: Avatarify lets users run realtime deepfakes on live video calls

#86

It would be interesting to see how far you could get using deepfakes as a method for video call compression. Train a model locally ahead of time and upload it to a server, then whenever you have a call scheduled the model is downloaded in advance by the other participants. Now, instead of having to send video data, you only have to send a representation of the facial movements so that the recipients can render it on…

Excellent idea and we'll surely be seeing something like this, there are AR apps that already map facial expressions to avatars. Downside could be some uncanny valley if the models are not very high quality. But if I had to make a prediction, I'd expect we'll get much more value from higher bandwidth, ultra high definition streaming and features like 3d cameras / virtual reality. I think we have a tendency to really…

> Downside could be some uncanny valley if the models are not very high quality.

That can be controlled, since these compression algorithms usually work by making a prediction and sending the difference between the prediction and the actual value.

That works both for lossless compression - where the difference is sent in full - and lossy as well - where only the most important part of the difference is sent.

Re: Avatarify lets users run realtime deepfakes on live video calls

#88
post #58

Earlier quoted context omitted.

Excellent idea and we'll surely be seeing something like this, there are AR apps that already map facial expressions to avatars. Downside could be some uncanny valley if the models are not very high quality. But if I had to make a prediction, I'd expect we'll get much more value from higher bandwidth, ultra high definition streaming and features like 3d cameras / virtual reality. I think we have a tendency to really…

> I'd expect we'll get much more value from higher bandwidth, ultra high definition streaming and features like 3d cameras / virtual reality. I think we have a tendency to really underestimate how important high definition is for human communication. Low latency is probably more important to me. Recently I seem to have a 3 second delay on many VC calls at work (and just for me it seems), and I either end up interrupt…

This can be helped with hand-raising (queue style) and a dedicated facilitator for each meeting.

Re: Avatarify lets users run realtime deepfakes on live video calls

#89

I am just waiting for someone to build a deep_nude_realtime_zoom plugin so I can finally tell people that we should take digital security, privacy and identity seriously.

Absolutely. I could do my video conference calls naked and blame the result on my new deep fake software. Hell, I can even do that now without having the software. My colleagues are nerdy enough to assume such software already exists

Re: Avatarify lets users run realtime deepfakes on live video calls

#90

It would be interesting to see how far you could get using deepfakes as a method for video call compression. Train a model locally ahead of time and upload it to a server, then whenever you have a call scheduled the model is downloaded in advance by the other participants. Now, instead of having to send video data, you only have to send a representation of the facial movements so that the recipients can render it on…

I think this is largely possible, and accuracy to a human is very different than MSE accuracy used in a traditional lossy compression algorithm.

To a human, for example, the exact pattern of every strand of hair isn't important at all -- all that matters is that the hairstyle and hair color stays the same.

The algorithm can also not worry about encoding and re-constructing skin blemishes because humans would possibly actually enjoy not having to put on makeup for a video call.

Post reply on HN