Live data from Hacker News

VideoGigaGAN: Towards detail-rich video super-resolution

videogigagan.github.io

101–110 of 243 posts

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#101

Earlier quoted context omitted.

if the nn is part of the codec, you can choose to only downscale the regions that get reconstructed correctly.

Why would you not let the NN work on the compressed data? That is actually where the information is.

that's like asking why you don't train a llm on gzipped text. the compressed data is much harder to reason about

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#102

Earlier quoted context omitted.

At 30fps, which is not high, that would mean chunks of less than 7 seconds. Doable but highly impractical to say the least.

You will probably have to have some overhang of time to get the state space to match enough to minimize flicker in between fragments.

You could probably mitigate this by using overlapping clips and fading between them. Pretty crude but could be close to unnoticeable, depending on how unstable the technique actually is.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#103
post #57

Earlier quoted context omitted.

A few years if not less. They will have huge budgets for compute and the makers of compute will be happy to absorb those budgets. Cloud production was already growing but this will continue to accelerate it imho

Wasn't Hollywood an early adopter of advanced AI video stuff, w.r.t. de-aging old famous actors?

Bingo. Except it looked like magic because the tech was so expensive and only available to them.

Limited access to the tech added some mystique to it too.

Just like digital cameras created a lot more average photographers, it pushed photography to a higher standard than just having access to expensive equipment.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#105
post #95

Is anyone else concerned at the societal effects of technology like this? In one of the examples they show a young girl. In the upscale example it's quite clearly hallucinating makeup and lipstick. I'm quite worried about tools like this perpetuating social norms even further.

Yes, but if you mention that here, you’ll get accused of wokeism. More seriously, though, yes, the thing you’re describing is exactly what the AI safety field is attempting to address.

> is exactly what the AI safety field is attempting to address

Is it though? I think it's pretty obvious to any neutral observer that this is not the case, at least judging based on recent examples (leading with the Gemini debacle).

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#106

Video quality seems really good, but limitations are quite restrictive "Our model encounters challenges when processing extremely long videos (e.g. 200 frames or more)". I'd say most videos in practice are longer than 200 frames, so lot more research is still needed.

Unless they can predict a 2 hour movie in 200 frames.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#107

Is anyone else concerned at the societal effects of technology like this? In one of the examples they show a young girl. In the upscale example it's quite clearly hallucinating makeup and lipstick. I'm quite worried about tools like this perpetuating social norms even further.

Aside your point: It does look like she is wearing lipstick tho, to me. More likely lip balm. Her (unaltered) lips have specular highlights on the tops that suggests they're wet or have lip balm to me. As for the makeup, not sure there. Here cheeks seem rosy in the original, and not sure what you're referring to beyond that. Perhaps her skin is too clear in the AI version, suggesting some type of foundation?

I know nothing of makeup tho, just describing my observations.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#108
post #95

Earlier quoted context omitted.

Yes, but if you mention that here, you’ll get accused of wokeism. More seriously, though, yes, the thing you’re describing is exactly what the AI safety field is attempting to address.

> is exactly what the AI safety field is attempting to address Is it though? I think it's pretty obvious to any neutral observer that this is not the case, at least judging based on recent examples (leading with the Gemini debacle).

Yes, avoiding creating societally-harmful content is what the Gemini "debacle" was attempting to do. It clearly had unintended effects (e.g: generating a black Thomas Jefferson), but when these became apparent, they apologized and tried to put up guard rails to keep those negative effects from happening.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#109
post #58

Earlier quoted context omitted.

At 24fps that's not even 10 seconds. Calling it extremely long is kinda defensive.

The average shot length in a modern movie is around 2.5 seconds (down from 12 seconds in 1930's). For animations it's around 15 seconds.

Huh, I thought this couldn't be true, but it is. The first time I noticed annoyingly fast cuts was World War Z, for me it was unwatchable with tons of shots around 1 second each.

Re: VideoGigaGAN: Towards detail-rich video super-resolution

#110
The video of the owl is a great example of doing a terrible job without the average Joe noticing.

The real owl has fine light/dark concentric circles on its face. The app turned it into gray because it does not see any sign of the circles. The real owl has streaks of spots. The app turned them into solid streaks because it saw no sign of spots. There's more where this came from, but basically only looks good to someone who has no idea what the owl should look like.

Post reply on HN