Earlier quoted context omitted.
if the nn is part of the codec, you can choose to only downscale the regions that get reconstructed correctly.
Why would you not let the NN work on the compressed data? That is actually where the information is.
VideoGigaGAN: Towards detail-rich video super-resolution
101–110 of 243 posts
Re: VideoGigaGAN: Towards detail-rich video super-resolution
#102Earlier quoted context omitted.
At 30fps, which is not high, that would mean chunks of less than 7 seconds. Doable but highly impractical to say the least.
You will probably have to have some overhang of time to get the state space to match enough to minimize flicker in between fragments.
Re: VideoGigaGAN: Towards detail-rich video super-resolution
#103Earlier quoted context omitted.
A few years if not less. They will have huge budgets for compute and the makers of compute will be happy to absorb those budgets. Cloud production was already growing but this will continue to accelerate it imho
Wasn't Hollywood an early adopter of advanced AI video stuff, w.r.t. de-aging old famous actors?
Limited access to the tech added some mystique to it too.
Just like digital cameras created a lot more average photographers, it pushed photography to a higher standard than just having access to expensive equipment.
Re: VideoGigaGAN: Towards detail-rich video super-resolution
#104Re: VideoGigaGAN: Towards detail-rich video super-resolution
#105Is anyone else concerned at the societal effects of technology like this? In one of the examples they show a young girl. In the upscale example it's quite clearly hallucinating makeup and lipstick. I'm quite worried about tools like this perpetuating social norms even further.
Yes, but if you mention that here, you’ll get accused of wokeism. More seriously, though, yes, the thing you’re describing is exactly what the AI safety field is attempting to address.
Is it though? I think it's pretty obvious to any neutral observer that this is not the case, at least judging based on recent examples (leading with the Gemini debacle).
Re: VideoGigaGAN: Towards detail-rich video super-resolution
#106Video quality seems really good, but limitations are quite restrictive "Our model encounters challenges when processing extremely long videos (e.g. 200 frames or more)". I'd say most videos in practice are longer than 200 frames, so lot more research is still needed.
Re: VideoGigaGAN: Towards detail-rich video super-resolution
#107Is anyone else concerned at the societal effects of technology like this? In one of the examples they show a young girl. In the upscale example it's quite clearly hallucinating makeup and lipstick. I'm quite worried about tools like this perpetuating social norms even further.
I know nothing of makeup tho, just describing my observations.
Re: VideoGigaGAN: Towards detail-rich video super-resolution
#108Earlier quoted context omitted.
Yes, but if you mention that here, you’ll get accused of wokeism. More seriously, though, yes, the thing you’re describing is exactly what the AI safety field is attempting to address.
> is exactly what the AI safety field is attempting to address Is it though? I think it's pretty obvious to any neutral observer that this is not the case, at least judging based on recent examples (leading with the Gemini debacle).
Re: VideoGigaGAN: Towards detail-rich video super-resolution
#109Earlier quoted context omitted.
At 24fps that's not even 10 seconds. Calling it extremely long is kinda defensive.
The average shot length in a modern movie is around 2.5 seconds (down from 12 seconds in 1930's). For animations it's around 15 seconds.
Re: VideoGigaGAN: Towards detail-rich video super-resolution
#110The real owl has fine light/dark concentric circles on its face. The app turned it into gray because it does not see any sign of the circles. The real owl has streaks of spots. The app turned them into solid streaks because it saw no sign of spots. There's more where this came from, but basically only looks good to someone who has no idea what the owl should look like.