Live data from Hacker News

VASA-1: Lifelike audio-driven talking faces generated in real time

microsoft.com

71–80 of 166 posts

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#71

I'm curious what is the reason for deepfake research and what the practical application is. Can someone explain the commercial need to take someones likeness and generate video content? If I was an a-list celebrity, I would give permission for coke to make a commercial with my likeness, provided I am allowed final approval of the finished ad? Do I have an avatar that attends my zoom work calls?

State disinformation and propaganda campaigns.

Corporate disinformation and propaganda campaigns.

Personal disinformation and propaganda campaigns.

Oh Brave New World, that has such fake people in it!

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#72

I'm curious what is the reason for deepfake research and what the practical application is. Can someone explain the commercial need to take someones likeness and generate video content? If I was an a-list celebrity, I would give permission for coke to make a commercial with my likeness, provided I am allowed final approval of the finished ad? Do I have an avatar that attends my zoom work calls?

Propaganda, political manipulation, narrative nudging, regular scams and advertising.

Even though most of those things are illegal you could just have foreign cat's paw firms do it. Maybe you fire them for "going to far" after the damage is done, assuming some even manages to connect the dots.

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#73

Why is this research being done? Is this some kind of arms race? The only purpose of this technology I can think of is getting spies to abuse others. Am I going to have to do AuthN and AuthZ on every phone call and zoom now?

On the other hand, if deepfaking becomes common enough that everyone stops trusting everything they read / see on the internet, it would be a net good against the spread of disinformation compared to today.

I don't see the extinction of trust through the introduction of garbage falsehoods to be a net good.

Believing that everything you eat is poisoned is no way to live. Believing that everything you see is a lie is also no way to live.

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#74

Anyone have any good ideas for how we're going to do politics now? Today a big ML model can do this and it's somewhat regulate-able, tomorrow people can do this on their contact-lens supercomputers and anyone can generate a video of anything. Is going back to personally knowing your local representative the only way? How will we vote for national candidates if nobody knows what they think or say?

> Today a big ML model can do this

Not that big:

https://github.com/Zejun-Yang/AniPortrait

https://huggingface.co/ZJYang/AniPortrait/tree/main

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#75

Anyone have any good ideas for how we're going to do politics now? Today a big ML model can do this and it's somewhat regulate-able, tomorrow people can do this on their contact-lens supercomputers and anyone can generate a video of anything. Is going back to personally knowing your local representative the only way? How will we vote for national candidates if nobody knows what they think or say?

Hyper targeted placement of generated content designed to entice you to donate to political campaigns and to vote. Perhaps leading to a point where entire video clips are generated for a single viewer. Politicians and political commentators will lease their likeness and voice out for targeted messaging to be generated using their likeness. Less reputable platforms will allow disinformation campaigns to spread.

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#76

Anyone have any good ideas for how we're going to do politics now? Today a big ML model can do this and it's somewhat regulate-able, tomorrow people can do this on their contact-lens supercomputers and anyone can generate a video of anything. Is going back to personally knowing your local representative the only way? How will we vote for national candidates if nobody knows what they think or say?

DNS? Might be that we need a radical (for some) change of viewpoint.

Just as there's no privacy on the internet, how about 'theres very little trust on the internet'. Assume everything not securely signed by a trusted party is false.

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#77

I'm curious what is the reason for deepfake research and what the practical application is. Can someone explain the commercial need to take someones likeness and generate video content? If I was an a-list celebrity, I would give permission for coke to make a commercial with my likeness, provided I am allowed final approval of the finished ad? Do I have an avatar that attends my zoom work calls?

[dead]

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#78
post #66

Anyone have any good ideas for how we're going to do politics now? Today a big ML model can do this and it's somewhat regulate-able, tomorrow people can do this on their contact-lens supercomputers and anyone can generate a video of anything. Is going back to personally knowing your local representative the only way? How will we vote for national candidates if nobody knows what they think or say?

People in my circles have been saying this for a few years now, and we've yet to see it happen. I've got my popcorn ready. But you can rest easy. Everyone just votes for the candidate their party picked, anyway.

It'll happen - deepfakes aren't good enough yet. But when they become ubiquitous and hard to spot, it'll be chaos until the average person is mentally inoculated against believing any video / anything on the internet.

I wonder if it's possible to digitally sign footage as it's captured? It'd be nice to have some share-able demonstrably true media.

Edit: I'm a centrist and I definitely would lean one way or the other based on who the options are (or who I think they are).

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#79
Despite vast investment in AI by VCs and vast numbers of startups in the field, these sort of things remain unavailable as simple consumer installable software.

Every second day HN has some post about some new amazing AI system. Never available to download run and use.

Why the vast investment and no startup selling consumer downloadable software to do it?

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#80

Earlier quoted context omitted.

Video games, entertainment, and avatars seems like the big ones.

If that is really the reason then this is insane and everyone involved should put their keyboards down and stop what they are doing. This would be as if we invented and sold nuclear weapons to dig out quarry mines faster. The inconvenience it saves us quickly disappears into the overwhelming shadow of the enormous harm now enabled.

> This would be as if we invented and sold nuclear weapons to dig out quarry mines faster.

”Project Plowshare was the overall United States program for the development of techniques to use nuclear explosives for peaceful construction purposes.”[0]

0: https://en.wikipedia.org/wiki/Project_Plowshare

Post reply on HN