Live data from Hacker News

Imagen Video: high definition video generation with diffusion models

imagen.research.google

51–60 of 500 posts

Re: Imagen Video: high definition video generation with diffusion models

#51

The ethical implications of this are huge. Paper does a good detailing of this. Very happy to see that the researchers are being cautious. edit: Just because it is cool to hate on AI ethics doesn't diminish the importance of using AI responsibly.

I feel stupid what are those ethical implications? It seems like just a cool technology to me.

Re: Imagen Video: high definition video generation with diffusion models

#52
post #49

Earlier quoted context omitted.

I like your optimism but OP's job is to take text instructions and turn them into video, for advertisements. If Google (who already control so much of the advertising space) can take text instructions and turn them into advertisements, what's left for OP to do here? Even if there's some additional editing required this seems like it will greatly reduce the hours an editor is needed. And it can probably iterate option…

Maybe OP's future involves being able to do their work 10x faster, while producing much higher quality results than people who have been given access to a generative AI model without first spending a decade+ learning what makes a good film clip. The optimistic view of all of this is that these tools will give people with skill and experience a massive productivity boost, allowing them to do the best work of their car…

No post body was provided.

Re: Imagen Video: high definition video generation with diffusion models

#53

I’m going to post an Ask HN about what am I supposed to do when I’m “disrupted”. I work in film / video / CG where the bread and butter is short form advertising for Youtube, Instagram and TV. It’s painfully obvious that in 1 year the job might be exceedingly more difficult than it is now.

What happened to volume of web and graphic designers when templates+wordpress hit them?

Re: Imagen Video: high definition video generation with diffusion models

#54

What everyone is missing is that these AI image/video generators lack _taste_. These tools just regurgitate a mishmash of images from it's training set, without any "feeling". What you're going to tell me that you can train them to have feeling? It's never going to happen.

"These tools just regurgitate a mishmash of images from it's training set"

I don't think that's a particularly useful mental model for how these work.

The models end up being a tiny fraction of the size of the training set - Stable Diffusion is just 4.3GB, it fits on a DVD!

So it's not a case of models pasting in bits of images they've seen - they genuinely do have a highly compressed concept of what a cactus looks like, which they can use to then render a cactus - but the thing they render is more of an average of every cactus they've seen rather than representing any single image that they were trained on.

But I agree with you on taste! This is why I'm most excited about what happens when a human with great taste gets to take control of these generative models and use them to create art that wouldn't be possible to create without them (or at least not possible to create within a short time-frame).

Re: Imagen Video: high definition video generation with diffusion models

#56

The ethical implications of this are huge. Paper does a good detailing of this. Very happy to see that the researchers are being cautious. edit: Just because it is cool to hate on AI ethics doesn't diminish the importance of using AI responsibly.

I feel stupid what are those ethical implications? It seems like just a cool technology to me.

Top two comments are creatives wondering about their future jobs. Ai ethicists have brought up concerns regarding intentional misuse like misinformation.

The technology is super cool. Cat is out of the bag. Just like we couldn't really make cryptography illegal, this stuff shouldn't be either. But I dislike how everyone is pretending that AI ethicists and others are completely unfounded just because it is popular to hate on them nowadays. Way too many people supported Y. Kilcher's antics.

The paper itself has more details.

Re: Imagen Video: high definition video generation with diffusion models

#57

What everyone is missing is that these AI image/video generators lack _taste_. These tools just regurgitate a mishmash of images from it's training set, without any "feeling". What you're going to tell me that you can train them to have feeling? It's never going to happen.

This isn't a very compelling argument. First of all, they aren't a "mish mash" in any real way, it's not like snippets of images exist inside of the model. Second of all, this is entirely subjective. Third of all, entirely inconsequential - if these models create 80% of the video we end up seeing, is it going to matter if you don't think it's a tasteful endeavour?

Re: Imagen Video: high definition video generation with diffusion models

#58
post #3

Probably only 6 months until we get this in stable diffusion format. Things are about to get nuts and awesome.

Emad (founder of Stability AI) has said they already have video model training underway, as well as text and audio. Exciting times.

Is this going to end up into a single model, where its trained on text and images and audio and videos and 3d models, and it can do anything to anything depending on what you ask of it? Feels like the cross-training would help yield stronger results.

Re: Imagen Video: high definition video generation with diffusion models

#59

I’m going to post an Ask HN about what am I supposed to do when I’m “disrupted”. I work in film / video / CG where the bread and butter is short form advertising for Youtube, Instagram and TV. It’s painfully obvious that in 1 year the job might be exceedingly more difficult than it is now.

What happened to volume of web and graphic designers when templates+wordpress hit them?

A lot of additional work, because the industry was growing like crazy in tandem.

Re: Imagen Video: high definition video generation with diffusion models

#60

This appears to understand and generate text much better. Hopefully just a few years to a prompt of "4k, widescreen render of this Star Trek: TNG episode".

At the rate this is going we are only a few years from generating a new TNG episode
Post reply on HN