Live data from Hacker News

Imagen Video: high definition video generation with diffusion models

imagen.research.google

61–70 of 500 posts

Re: Imagen Video: high definition video generation with diffusion models

#61
post #41

The total number of hyperparameters (sum of all the model blocks) is 16.25B, which is large but less than expected.

I assume you meant just "parameters" since "hyperparameters" has a specific alternate meaning? Sorry for the pedantry lol.

The AI world can't decide either.

Re: Imagen Video: high definition video generation with diffusion models

#62
post #29

I’m going to post an Ask HN about what am I supposed to do when I’m “disrupted”. I work in film / video / CG where the bread and butter is short form advertising for Youtube, Instagram and TV. It’s painfully obvious that in 1 year the job might be exceedingly more difficult than it is now.

When you animate a horse, does it have 5 legs with weird backwards joints? If not, your job is probably safe for now.

Think about where this stuff was 2 years ago and then think about where it will be 2 years from now.

Re: Imagen Video: high definition video generation with diffusion models

#63

"We have decided not to release the Imagen Video model or its source code until these concerns are mitigated" Okay then why even post it in the first place? What exactly is Google going to do with this model?

They're going to 1) rent it out as a paid API and/or 2) let you use it to create ads on Google platforms like YouTube, perhaps customized to the individual user

Re: Imagen Video: high definition video generation with diffusion models

#64
post #20

I’m going to post an Ask HN about what am I supposed to do when I’m “disrupted”. I work in film / video / CG where the bread and butter is short form advertising for Youtube, Instagram and TV. It’s painfully obvious that in 1 year the job might be exceedingly more difficult than it is now.

Whatever insights and expertize you've gained up until now can probably be used to gain enough of a competitive advantage in this future industry to be employed. I doubt the people that will spend their time on this professionally will be former coders etc. (I've seen the stable diffusion outputs that coders will tweet. It's a good illustration that taste is still hugely important.)

I think there will be tons of jobs that resemble software development for proper, quick high quality generation of video/images.

That being said, it’s possible that it won’t pay anywhere near what you’re used to. Either way, it will probably be a solid decade before you’ve really felt the pain for disruption. MP3s, which were a far more straightforward path to disruption took at least that long from conception.

Re: Imagen Video: high definition video generation with diffusion models

#65

Earlier quoted context omitted.

Emad (founder of Stability AI) has said they already have video model training underway, as well as text and audio. Exciting times.

Is this going to end up into a single model, where its trained on text and images and audio and videos and 3d models, and it can do anything to anything depending on what you ask of it? Feels like the cross-training would help yield stronger results.

These diffusion models are using a frozen text encoder (e.g. CLIP for Stable Diffusion, T5 for Imagen), which can be used in other applications.

StabilityAI trained a new/better CLIP for the purpose of better Stable Diffusions.

Re: Imagen Video: high definition video generation with diffusion models

#66

I’m going to post an Ask HN about what am I supposed to do when I’m “disrupted”. I work in film / video / CG where the bread and butter is short form advertising for Youtube, Instagram and TV. It’s painfully obvious that in 1 year the job might be exceedingly more difficult than it is now.

Start making content and charging for it. You no longer need institutional capital to make a Disney- or Pixar-like experience.

Small creators will win under this new regime of tools. It's a democratizing force.

Re: Imagen Video: high definition video generation with diffusion models

#68
post #20

Earlier quoted context omitted.

Whatever insights and expertize you've gained up until now can probably be used to gain enough of a competitive advantage in this future industry to be employed. I doubt the people that will spend their time on this professionally will be former coders etc. (I've seen the stable diffusion outputs that coders will tweet. It's a good illustration that taste is still hugely important.)

I think there will be tons of jobs that resemble software development for proper, quick high quality generation of video/images. That being said, it’s possible that it won’t pay anywhere near what you’re used to. Either way, it will probably be a solid decade before you’ve really felt the pain for disruption. MP3s, which were a far more straightforward path to disruption took at least that long from conception.

> That being said, it’s possible that it won’t pay anywhere near what you’re used to.

Also won't nearly require the amount of work it used to.

Re: Imagen Video: high definition video generation with diffusion models

#69

Earlier quoted context omitted.

I feel stupid what are those ethical implications? It seems like just a cool technology to me.

Top two comments are creatives wondering about their future jobs. Ai ethicists have brought up concerns regarding intentional misuse like misinformation. The technology is super cool. Cat is out of the bag. Just like we couldn't really make cryptography illegal, this stuff shouldn't be either. But I dislike how everyone is pretending that AI ethicists and others are completely unfounded just because it is popular to…

It’s impressive that the small videos are generated this way but the videos themselves are obviously ML generated as they are distorted, a lot like the other art, you can kinda tell it’s the computer. I’m not seeing the ethical issues. I mean cameras disrupted lots of jobs. In general that’s what all technology does everyday. What’s different about this technology?

Re: Imagen Video: high definition video generation with diffusion models

#70

"We have decided not to release the Imagen Video model or its source code until these concerns are mitigated" Okay then why even post it in the first place? What exactly is Google going to do with this model?

Ask the "AI Ethicists". They have to justify their salaries in some way or another. Or maybe Google is using "Responsible AI" as an excuse to minimize competitors when they release their own Imagen Video as a Service API in Google Cloud. It's quite strange when the "ethical" thing to do is to not publicly release your research, put it behind a highly restrictive API and charge a high price for it ($0.02 per 1k tokens…

This doesn’t really prevent competition though, the research paper is enough to recreate it. It does make recreation more expensive, but maybe that leaves you with a motivation to get paid for doing it.
Post reply on HN