Live data from Hacker News

Imagen Video: high definition video generation with diffusion models

imagen.research.google

111–120 of 500 posts

Re: Imagen Video: high definition video generation with diffusion models

#111
post #102

These are baby steps towards what I think will be the eventual "disruption" to the film and tv industry. Directors will simply be able to write a script/prompt long enough and detailed enough for something like Imagen (or it's successors) to convert into a feature-length show. Certainly we're very, very far away from that level of cinematic detail and crispness. But I believe that is where this leads... complete with…

We thought creative jobs were going to be the last thing AI replaces, now it's among the first.

What's next that may be counterintuitive?

Re: Imagen Video: high definition video generation with diffusion models

#112
post #84

I’m going to post an Ask HN about what am I supposed to do when I’m “disrupted”. I work in film / video / CG where the bread and butter is short form advertising for Youtube, Instagram and TV. It’s painfully obvious that in 1 year the job might be exceedingly more difficult than it is now.

There's a huge gap between "that's pretty cool" and a feature length film. People want to create specific stories with specific scenes in specific places that look a specific way. A "Couple kissing in the rain " prompt isn't going to produce something people are going to pay to see. It's more likely that you're still going to be filming/editing/animating but will have an AI layer on top that produces extra effects or…

I actually think it's the opposite, AI will probably be writing the stories and humans might occasionally film a few scenes. ~95% of TV shows and movies are cookie-cutter content, with cookie-cutter acting and production values, with the same hooks and the same tropes regurgitated over and over again. Heck they can't even figure out how to make new IP so they keep making reruns of the same old stuff like Star Wars, Marvel, etc, and people eat it right up. There's nothing better at figuring out how to maximize profit and hook people to watch another episode than a good algorithm.

Re: Imagen Video: high definition video generation with diffusion models

#113
Google continues to blow my mind with these models, but I think their ethics strategy is totally misguided and will result in them failing to capture this market. The original Google Search gave similarly never-before-seen capabilities to people, and you could use it for good or bad - Google did not seem to have any ethical concerns around, for example, letting children use their product and come across NSFW content (as a kid who grew up with Google you can trust me on this).

But now with these models they have such a ridiculously heavy handed approach to the ethics and morals. You can't type any prompt that's "unsafe", you can't generate images of people, there are so many stupid limitations that the product is practically useless other than niche scenarios, because Google thinks it knows better than you and needs to control what you are allowed to use the tech for.

Meanwhile other open source models like Stable Diffusion have no such restrictions and are already publicly available. I'd expect this pattern to continue under Google's current ideological leadership - Google comes up with innovative revolutionary model, nobody gets to use it because "safety", and then some scrappy startup comes along, copies the tech, and eats Google's lunch.

Google: stop being such a scared, risk averse company. Release the model to the public, and change the world once more. You're never going to revolutionize anything if you continue to cower behind "safety" and your heavy handed moralizing.

Re: Imagen Video: high definition video generation with diffusion models

#114

The ethical implications of this are huge. Paper does a good detailing of this. Very happy to see that the researchers are being cautious. edit: Just because it is cool to hate on AI ethics doesn't diminish the importance of using AI responsibly.

AI Ethics is a joke. It's literally Philip Morris funding research into the risks of smoking and concluding the worst that can happen to you is burning your hand.

Re: Imagen Video: high definition video generation with diffusion models

#115
I’ll be honest, as someone who worked in the film industry for a decade, this thread is depressing.

It’s not the technology, it’s all the people in these comments who have never worked in the industry clamouring for its demise.

One could brush it off as tech heads being over exuberant, but it’s the lack of understanding of how much fine control goes into each and every shot of a film that is depressing.

If I, as a creative, made a statement that security or programming is easy while pointing to GitHub Copilot, these same people would get defensive about it because they’d see where the deficiencies are.

However because they’re so distanced from the creative process, they don’t see how big a jump it is from where this or stage diffusion is to where even a medium or high tier artist are.

You don’t see how much choice goes into each stroke, or wrinkle fold , how much choice goes into subtle movements. More importantly you don’t see the iterations or emotional storytelling choices even in a character drawing or pose. You don’t see the combined decades, even centuries of experience, that go into making the shot and then seeing where you can make it better based on intangibles

So yeah this technology is cool, but I think people saying this will disrupt industries with vigour need to immerse themselves first before they comment as outsiders.

Re: Imagen Video: high definition video generation with diffusion models

#116
post #103
post #102

These are baby steps towards what I think will be the eventual "disruption" to the film and tv industry. Directors will simply be able to write a script/prompt long enough and detailed enough for something like Imagen (or it's successors) to convert into a feature-length show. Certainly we're very, very far away from that level of cinematic detail and crispness. But I believe that is where this leads... complete with…

I really doubt you’d be able to have the fine grained control that most high end creatives want with any of these diffusion models, let alone the ability to convey specific emotions. At that point, we’d have reached some kind of AI singularity and the disruption would be everywhere not just in the creative sphere

There's no doubt that it's only a matter of time.

Like bloggers had the opportunity to compete with newspapers, the ability to generate videos will allow to compete with movies/marvel/netflix/disney & company.

Eventually, only high quality content will justify the need to pay for a ticket or a subscription, and there's going to be a lot of free content to watch, with 1000x more people able to publish their ideas, as many have been doing with code on github for a while now, disrupting the concept of closed source code.

Re: Imagen Video: high definition video generation with diffusion models

#117
post #102

These are baby steps towards what I think will be the eventual "disruption" to the film and tv industry. Directors will simply be able to write a script/prompt long enough and detailed enough for something like Imagen (or it's successors) to convert into a feature-length show. Certainly we're very, very far away from that level of cinematic detail and crispness. But I believe that is where this leads... complete with…

There's already a surplus of video and an apparent lack of _quality_ video. This might be enough to get folks to shut the TV off completely.

Re: Imagen Video: high definition video generation with diffusion models

#118

Google continues to blow my mind with these models, but I think their ethics strategy is totally misguided and will result in them failing to capture this market. The original Google Search gave similarly never-before-seen capabilities to people, and you could use it for good or bad - Google did not seem to have any ethical concerns around, for example, letting children use their product and come across NSFW content…

[deleted]

Re: Imagen Video: high definition video generation with diffusion models

#119
post #88
post #66

Earlier quoted context omitted.

Start making content and charging for it. You no longer need institutional capital to make a Disney- or Pixar-like experience. Small creators will win under this new regime of tools. It's a democratizing force.

> It's a democratizing force. I'm wondering why the open source community doesn't get this. So many voices were raised against Codex. Now artists against Diffusion models. But the model itself is a distillation of everything we created, it can compactly encode it and recreate it in any shape and form we desire. That means everyone gets to benefit, all skills are available for everyone, all tailored to our needs.

> all skills are available for everyone

Exactly this!

We no longer have to pay the 10,000 hours to specialize.

The opportunity cost to choose our skill sets is huge. In the future, we won't have to contend with that horrible choice anymore. Anyone will be able to paint, play the piano, act, code, and more.

Post reply on HN