Live data from Hacker News

Nvidia Vid2vid: High-resolution photorealistic video-to-video translation

github.com

61–70 of 129 posts

Re: Nvidia Vid2vid: High-resolution photorealistic video-to-video translation

#62
post #59
post #54

Earlier quoted context omitted.

I don't know why you're talking about QR codes and transcripting/encrypting the audio when you can just sign the file? gpg --sign video.mp4

I assume they are talking about signing/watermarking a video in a way that survives video encoding/lossy transmission. A QR code probably wouldn’t work for various reasons but is an easy mental analogy to think of.

For embedding, you should just put it in the metadata. Encoding it in the video itself... I don't really see the point.

Re: Nvidia Vid2vid: High-resolution photorealistic video-to-video translation

#63

Earlier quoted context omitted.

It could also lead to crimes like Blackmail becoming extinct. It would be hard to hold incriminating recordings of anyone over them if near-perfect audio and video synthesis was common. Especially for public figures with lots training data available.

Yeah... The problem is try explaining deepfakes to your significant other when they are randomly sent what looks like a video of you cheating on them. Sure it’s possible but not likely.

It's all a matter of cultural awareness. Everyone now thinks when seeing an unlikely photo -- "Photoshop?"

This stuff even has the catchy name "deep fake".

Re: Nvidia Vid2vid: High-resolution photorealistic video-to-video translation

#64

Earlier quoted context omitted.

The vast bulk of Adobe's advantage is in UX, not technical algorithms. Which makes perfect sense because that tends to be the case with most F/OSS software—technically brilliant but with an face only a programmer could love. Yes, Adobe do have some remarkable algorithms that would be difficult to replicate (e.g. heal brush and content aware fill) but these are a small minority of Adobe's software advantage. The one t…

Do you use the Astute plugins for AI, or the native pen tool? Affinity feels different, maybe less precise, but the functionality was way better compared with the native Adobe tool. You should try Figmas pen tool, I like it.

My uses are relatively trivial- logo design, SVG generation, basic layout work and PDF tinkering. My main need is fine control of beziers with auto-guides to ensure consistency.

Re: Nvidia Vid2vid: High-resolution photorealistic video-to-video translation

#65

Earlier quoted context omitted.

Also, map images pulled from facebook to the bodies of pornstars. The creepiness and invasion of ... personal image(?) this enables is horrifying.

That's already happened. Search for deep fakes. We already have face substitution in videos which is working surprisingly well sometimes. There's been r/deepfakes where around 30% of the content was porn with swapped faces. It was banned though.

It was pretty limited though. For split seconds videos seemed real but the seams soon enough showed.

Fake celebrity porn has been on the internet since 1996 at the very least. It's always been crummy; but porn in general requires a thick suspension of disbelief and an intense focus in a partial object of desire (what Lacan calls the objet petit a) that blurs everything else.

Re: Nvidia Vid2vid: High-resolution photorealistic video-to-video translation

#66

I'm probably being captain obvious here, but if this is what's being released for free, I wonder how much better a polished commercial version does, and when we reach the point where we can't trust anything we see anymore. It doesn't even have to be super perfect, even reaching the point where it takes experts about two weeks to determine if something's real or not might already be long enough to do great damage. Fro…

It could also lead to crimes like Blackmail becoming extinct. It would be hard to hold incriminating recordings of anyone over them if near-perfect audio and video synthesis was common. Especially for public figures with lots training data available.

DRM will be pushed hard, starting from video/audio acquisition, perhaps assisted by blockchain to keep footage verified at all processing steps.

Re: Nvidia Vid2vid: High-resolution photorealistic video-to-video translation

#67

Some applications for this kind of tech: - Porn, yeah, first application you can think of, there are already some startups doing it. - Doubling actors, and applied to sound, maybe you could translate from one language to another but kind of keeping the accent and tone. - Propaganda and misinformation. Now you can get your enemy to say and do whatever you want, on video. - Photo-realistic games. Create a rough 3D mode…

A company that used this tech to actually change the lip movements of actors for each translation would stand to make a lot of money right now.

Re: Nvidia Vid2vid: High-resolution photorealistic video-to-video translation

#68
post #42

Earlier quoted context omitted.

MS Office is still miles ahead of any OSS “alternative“, despite the many glaring bugs.

MS Office is still miles ahead in the benchmark of opening it's own proprietary files.

There is no competitor, proprietary or open, that comes close to Excel. It's been relentlessly, extensively polished for years and years, and keeps gaining new features every year.

And this sticking to the spreadsheet concept, which is very limiting.

---

Contrast for example Tableau -- it's a great idea and generated a lot of enthusiasm for a while, but never quite took off as an office package one needs to have. The normal awkwardness of its first versions is still there; they don't have the deep that the Excel team has.

Re: Nvidia Vid2vid: High-resolution photorealistic video-to-video translation

#69
post #66

Earlier quoted context omitted.

It could also lead to crimes like Blackmail becoming extinct. It would be hard to hold incriminating recordings of anyone over them if near-perfect audio and video synthesis was common. Especially for public figures with lots training data available.

DRM will be pushed hard, starting from video/audio acquisition, perhaps assisted by blockchain to keep footage verified at all processing steps.

Don't you think that a blockchain that works for anything other than a rather useless currency should be created before suggesting one for such a use? I see comments all the time about how we should use blockchain for this and that, and yet far simpler uses for blockchain haven't yet worked out.

Re: Nvidia Vid2vid: High-resolution photorealistic video-to-video translation

#70

I'm probably being captain obvious here, but if this is what's being released for free, I wonder how much better a polished commercial version does, and when we reach the point where we can't trust anything we see anymore. It doesn't even have to be super perfect, even reaching the point where it takes experts about two weeks to determine if something's real or not might already be long enough to do great damage. Fro…

Yeah, the Tweet I found this from had a similar sentiment:

https://twitter.com/PiratePartyINT/status/104296466807811686...

"Starting now, we cannot trust video or audio evidence. The ramifications for our legal & political systems will not be known for many years"

Post reply on HN