Live data from Hacker News

VideoGigaGAN: Towards detail-rich video super-resolution

videogigagan.github.io

121–130 of 243 posts

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#121
post #95

Is anyone else concerned at the societal effects of technology like this? In one of the examples they show a young girl. In the upscale example it's quite clearly hallucinating makeup and lipstick. I'm quite worried about tools like this perpetuating social norms even further.

Yes, but if you mention that here, you’ll get accused of wokeism. More seriously, though, yes, the thing you’re describing is exactly what the AI safety field is attempting to address.

[deleted]

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#122

Video quality seems really good, but limitations are quite restrictive "Our model encounters challenges when processing extremely long videos (e.g. 200 frames or more)". I'd say most videos in practice are longer than 200 frames, so lot more research is still needed.

Well there goes my dreams of making my own Deep Space Nine remaster from DVDs.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#123
post #96

Is anyone else concerned at the societal effects of technology like this? In one of the examples they show a young girl. In the upscale example it's quite clearly hallucinating makeup and lipstick. I'm quite worried about tools like this perpetuating social norms even further.

I don't know, it's a mirror, right? It's up to us to change really. Besides, failures like the one you point out make subtle stereotypes and biases more conspicuous, which could be a good thing.

Precisely: tools don't have morality. We have to engage in political and social struggle to make our conditions better. These tools can help but they certainly wont do it for us, nor will they be the reason why things go bad.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#124
post #109

Earlier quoted context omitted.

The average shot length in a modern movie is around 2.5 seconds (down from 12 seconds in 1930's). For animations it's around 15 seconds.

Huh, I thought this couldn't be true, but it is. The first time I noticed annoyingly fast cuts was World War Z, for me it was unwatchable with tons of shots around 1 second each.

Yeah, the average may also be getting driven (e: down) by the basketball scene in Catwoman

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#125
post #96

Is anyone else concerned at the societal effects of technology like this? In one of the examples they show a young girl. In the upscale example it's quite clearly hallucinating makeup and lipstick. I'm quite worried about tools like this perpetuating social norms even further.

I don't know, it's a mirror, right? It's up to us to change really. Besides, failures like the one you point out make subtle stereotypes and biases more conspicuous, which could be a good thing.

It's interesting that the output of the genAI will inevitably get fed into itself. Both directly and indirectly by influencing humans who generate content that goes back into the machine. How long will the feedback loop take to output content reflecting new trends? How much new content is needed to be reflected in the output in a meaningful way. Can more recent content be weighted more heavily? Such interesting stuff!

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#126
I'm curious as to how well this works when upscaling from 1080p to 4K or 4K to 8K.

Their 128x128 to 1024x1024 upscales are very impressive, but I find the real artifacts and weirdness are created when AI tries to upscale an already relatively high definition image.

I find it goes haywire, adding ghosting, swirling, banded shadowing, etc as it whirlwinds into hallucinations from too much source data since the model is often trained to work with really small/compressed video into an "almost HD" video.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#127

Earlier quoted context omitted.

The average shot length in a modern movie is around 2.5 seconds (down from 12 seconds in 1930's). For animations it's around 15 seconds.

The textures of objects need to maintain consistency across much larger time frames, especially at 4k where you can see the pores on someone's face in a closeup.

I'm sure if you really want to burn money on compute you can do some smart windowing in the processing and use it on overlapping chunks and do an OK job.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#128

Wonder how long until Hollywood CGI shops have these types of models running as part of their post-production pipeline. Big blockbusters often release with ridiculously broken CGI due to crunch (Black Panther's third act was notorious for looking like a retro video-game), adding some extra generative polish in those cases is a no-brainer.

Once AI tech gets fully integrated entire Hollywood rendering pipeline will go from rendering to diffusing

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#129
I wonder if you could specialise a model by training it on a whole movie or TV series, so that instead of hallucinating from generic images, the model generates things it has seen closer-up in other parts of the movie.

You'd have to train it to go from a reduced resolution to the original resolution, then apply that to small parts of the screen at the original resolution to get an enhanced resolution, then stitch the parts together.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#130

I am personally much more interested in frame rate upscalers. A proper 60Hz just looks much better then anything. Also would really, really like to see a proper 60Hz animate upscale. Anything in that space just sucks. But when in the rare cases it works it really looks next level.

Have you tried DAIN?
Post reply on HN