Live data from Hacker News

One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing

nvidia-research-mingyuliu.com

21–30 of 52 posts

Re: One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing

#21
post #7

This makes a request to server to get the result back. Hacker News hug of death has already happened. I wish this was deployable to browsers so it was fully stand alone.

I wish they had done this client-side with e.g. Tensorflow.js. Would have been much more fun to play with.

Re: One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing

#22
Inviting terms as usual:

When you upload, submit, store, send or receive User Content to or through the NVIDIA Research AI Playground, you give NVIDIA (and parties NVIDIA works with, including its affiliates, suppliers and customers) a worldwide license to use (including without limitation for neural network training), host, store, reproduce, modify, create derivative works (such as those resulting from translations, adaptations or other changes), communicate, publish, publicly perform, publicly display and distribute such User Content. The rights you grant in this license are for the limited purpose of operating, promoting, and improving the NVIDIA Research AI Playground and content available to all users, and to develop new NVIDIA offerings. This license continues even if you stop using the NVIDIA Research AI Playground. The NVIDIA Research AI Playground may offer you ways to access, download, and remove content that has been provided, but make sure to keep your own back-up copies of your User Content. Also, the scope of services is limited and not all content in all formats can be loaded in the NVIDIA Research AI Playground.

Re: One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing

#23
post #19
post #12

Earlier quoted context omitted.

I like faces. Why deny a key part of being human?

If you don't know why one might want this, you might be a white male, or at least white, or at least male, or at least a member of the majority race in your locality. I've definitely wanted this on a few occasions, to avoid being discriminated.

So following this logic train, the only actual faces one may see would be white males. Everyone else gets a mask?

Re: One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing

#24
post #23
post #19

Earlier quoted context omitted.

If you don't know why one might want this, you might be a white male, or at least white, or at least male, or at least a member of the majority race in your locality. I've definitely wanted this on a few occasions, to avoid being discriminated.

So following this logic train, the only actual faces one may see would be white males. Everyone else gets a mask?

Maybe? But I'm going to get what I wanted.

Re: One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing

#25
post #24
post #23

Earlier quoted context omitted.

So following this logic train, the only actual faces one may see would be white males. Everyone else gets a mask?

Maybe? But I'm going to get what I wanted.

Would that then just be a proxy for not being a white male and present the same problems? I am not sure what this solves to prevent discrimination.

Re: One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing

#26
post #5
post #3

Wow, this works pretty well. Makes me think of that chapter in Infinite Jest where videoconferencing gets popular, until people start using "optimized" computer-rendered images instead of showing their actual faces, at which point everyone goes back to audio-only.

I wouldn't mind video conferencing with computer generated avatars. I don't video conference to know what the other person looks like, and in fact knowing what they look like just creates lots of unnecessary bias. I do it for the cues from their gestures, facial expressions, the direction they are looking, etc. With a good tracking setup that works perfectly well today with digital avatars.

There is a crappy movie called "Surrogates" that goes down this path.

Re: One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing

#28
post #25
post #24

Earlier quoted context omitted.

Maybe? But I'm going to get what I wanted.

Would that then just be a proxy for not being a white male and present the same problems? I am not sure what this solves to prevent discrimination.

It doesn't prevent discrimination, it just allows me to do what I want to do in the short term, e.g. raise funding for startups or get dream jobs or whatever.

I mean if a VC hands me a term sheet or a hiring manager hands me a job and the ONLY thing I misrepresented is my face, I don't think they have any ethical grounds to retract their offer.

Solving the discrimination problem is another matter, and will take years if not decades, my dreams cannot wait for that.

Re: One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing

#29
post #24
post #23

Earlier quoted context omitted.

So following this logic train, the only actual faces one may see would be white males. Everyone else gets a mask?

Maybe? But I'm going to get what I wanted.

Looking white male, isn't enough if you go down that road. Accents, tone, pitch of the voice would also need to match

Re: One-Shot Free-View Neural Talking-Head Synthesis for Video Conferencing

#30
post #29
post #24

Earlier quoted context omitted.

Maybe? But I'm going to get what I wanted.

Looking white male, isn't enough if you go down that road. Accents, tone, pitch of the voice would also need to match

That's much easier if you've spent enough time in the US and basically "sound white" (whatever that means) but don't look white.

But I imagine accent change could be another subject of a future deep learning project.

Post reply on HN