Live data from Hacker News

VASA-1: Lifelike audio-driven talking faces generated in real time

microsoft.com

11–20 of 166 posts

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#11
This is absolutely crazy. And it'll only get better from here. Imagine "VASA-9" or whatever.

I thought deepfakes were still quite a bit away but after this I will have to be way more careful online. It's not far from behind something that can show up in your "YouTube shorts" feed and trick you if you didn't already know it was AI.

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#13

“We have no plans to release an online demo, API, product, additional implementation details, or any related offerings until we are certain that the technology will be used responsibly and in accordance with proper regulations.”

/s it doesn’t have the phrase LLM in the title

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#14

“We have no plans to release an online demo, API, product, additional implementation details, or any related offerings until we are certain that the technology will be used responsibly and in accordance with proper regulations.”

> until we are certain that the technology will be used responsibly ...

That's basically "never" then, so we'll see how long they hold out.

Scammers are already using the existing voice/image/video generation apparently fairly successfully. :(

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#17

“We have no plans to release an online demo, API, product, additional implementation details, or any related offerings until we are certain that the technology will be used responsibly and in accordance with proper regulations.”

money will change that

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#19
post #11

This is absolutely crazy. And it'll only get better from here. Imagine "VASA-9" or whatever. I thought deepfakes were still quite a bit away but after this I will have to be way more careful online. It's not far from behind something that can show up in your "YouTube shorts" feed and trick you if you didn't already know it was AI.

This is good but nowhere as good as EMO https://humanaigc.github.io/emote-portrait-alive/ (https://news.ycombinator.com/item?id=39533326)

This one has too much movement and looks eerie/robotic/uncanny valley. While EMO looks just perfect.

Re: VASA-1: Lifelike audio-driven talking faces generated in real time

#20

“We have no plans to release an online demo, API, product, additional implementation details, or any related offerings until we are certain that the technology will be used responsibly and in accordance with proper regulations.”

> until we are certain that the technology will be used responsibly ... That's basically "never" then, so we'll see how long they hold out. Scammers are already using the existing voice/image/video generation apparently fairly successfully. :(

Eventually someone will implement one of these really good recent ones as open source and then it will be on replicate etc. right now the open source ones like SadTalker and Video Retalking are not live and are unconvincing.
Post reply on HN