Live data from Hacker News

Imagen Video: high definition video generation with diffusion models

imagen.research.google

101–110 of 500 posts

Re: Imagen Video: high definition video generation with diffusion models

#101
post #66

I’m going to post an Ask HN about what am I supposed to do when I’m “disrupted”. I work in film / video / CG where the bread and butter is short form advertising for Youtube, Instagram and TV. It’s painfully obvious that in 1 year the job might be exceedingly more difficult than it is now.

Start making content and charging for it. You no longer need institutional capital to make a Disney- or Pixar-like experience. Small creators will win under this new regime of tools. It's a democratizing force.

Outcome uncertain. Why would I need to buy content when I can generate my own with a local GPU?

Eventually the data model will be abstracted into deterministic code using a seed value; think implications of E=mc^2 being unpacked. The only “data” to download will be the source.

And the real world politics have not gone anywhere; none of us own the machines that produce the machines to run this. They could just sell locked down devices that will only iterate on their data structures.

There is no certainty “this time” we’ll pop “the grand illusion.”

Re: Imagen Video: high definition video generation with diffusion models

#102
These are baby steps towards what I think will be the eventual "disruption" to the film and tv industry. Directors will simply be able to write a script/prompt long enough and detailed enough for something like Imagen (or it's successors) to convert into a feature-length show.

Certainly we're very, very far away from that level of cinematic detail and crispness. But I believe that is where this leads... complete with AI actors (or real ones deep faked throughout the show).

For a while I thought "The Volume" was going to be the disruption to the industry. Now I think AI like this will eventually take it over.

https://www.comingsoon.net/movies/features/1225599-the-volum...

The main motivation will be production costs and time for studios, of which The Volume is already showing huge gains for Disney/ILM (just look at how much new star wars content has popped up within a matter of a few years). But i'm unsure if Disney has patented this tech and workflow and if other studios will be able to leverage it.

Regardless, AI/software will eat the world, and this will be one more step towards it. Exciting stuff.

Re: Imagen Video: high definition video generation with diffusion models

#103
post #102

These are baby steps towards what I think will be the eventual "disruption" to the film and tv industry. Directors will simply be able to write a script/prompt long enough and detailed enough for something like Imagen (or it's successors) to convert into a feature-length show. Certainly we're very, very far away from that level of cinematic detail and crispness. But I believe that is where this leads... complete with…

I really doubt you’d be able to have the fine grained control that most high end creatives want with any of these diffusion models, let alone the ability to convey specific emotions.

At that point, we’d have reached some kind of AI singularity and the disruption would be everywhere not just in the creative sphere

Re: Imagen Video: high definition video generation with diffusion models

#104
post #102

These are baby steps towards what I think will be the eventual "disruption" to the film and tv industry. Directors will simply be able to write a script/prompt long enough and detailed enough for something like Imagen (or it's successors) to convert into a feature-length show. Certainly we're very, very far away from that level of cinematic detail and crispness. But I believe that is where this leads... complete with…

AI will also be able to fill in dialog, plot points, etc.

Re: Imagen Video: high definition video generation with diffusion models

#105
post #102

These are baby steps towards what I think will be the eventual "disruption" to the film and tv industry. Directors will simply be able to write a script/prompt long enough and detailed enough for something like Imagen (or it's successors) to convert into a feature-length show. Certainly we're very, very far away from that level of cinematic detail and crispness. But I believe that is where this leads... complete with…

I think long-term, yes. If you include the whole multimediosphere of 2D inputs and the wealth of 3D engine magickry, yes.

How long? Could be decades. But ultimately, yes.

Re: Imagen Video: high definition video generation with diffusion models

#106

"We have decided not to release the Imagen Video model or its source code until these concerns are mitigated" Okay then why even post it in the first place? What exactly is Google going to do with this model?

This whole holier-than-thou moralizing strikes me as trying to steer the conversation away from the real issue, which came into spotlight with Stable Diffusion - one of authorship/violating the IP rights of artists, who now have come down in force against their would be tech overlords who are in the process or repackaging and reselling their work.

This forced ideological posturing of 'if we give it to the plebes, they are going to generate something naughty with it' masks the somehow more cynically evil take of big tech, who are essentially taking the entire creative output of humanity and reselling it as their own, piecemeal.

Additionally I think the Dalle vs. Stable Diffusion comparison highlights the true masters of these people (or at least the ones they dare not cross) - corporations with powerful IP lawyers. Just ask Dalle to generate a picture with Mickey Mouse - it won't be able to do it.

Re: Imagen Video: high definition video generation with diffusion models

#107
post #97

Earlier quoted context omitted.

How long do you think until the horse looks perfect? 12 months? 5 years? I’m still 30 and I don’t see how my industry won’t be entirely disrupted by this within the next decade. And that’s my optimistic projection. It could be we have amazing output in 24 months.

It's not about random short clips - imagine introducing a character like Mickey Mouse and reusing him everywhere with the same character - my guess is it's going to take a while until "transfer" like that will work reliably.

Dreambooth and Texual inversion is already here, and it's been just over a month since Stable Diffusion was released, so I'd bet on sooner rather than later.

https://github.com/XavierXiao/Dreambooth-Stable-Diffusion

https://textual-inversion.github.io/

Re: Imagen Video: high definition video generation with diffusion models

#109
post #103
post #102

These are baby steps towards what I think will be the eventual "disruption" to the film and tv industry. Directors will simply be able to write a script/prompt long enough and detailed enough for something like Imagen (or it's successors) to convert into a feature-length show. Certainly we're very, very far away from that level of cinematic detail and crispness. But I believe that is where this leads... complete with…

I really doubt you’d be able to have the fine grained control that most high end creatives want with any of these diffusion models, let alone the ability to convey specific emotions. At that point, we’d have reached some kind of AI singularity and the disruption would be everywhere not just in the creative sphere

[deleted]

Re: Imagen Video: high definition video generation with diffusion models

#110
post #34

I feel like in a not so far future, all this will be generalized into "generate new from all the existing". And at some point later, "all the existing" will be corrupted by the integrated "new" at it will all be chaos. I'm joking, it will be fun all along. :)

It's true, how will future AI train when the training datasets are themselves filled with AI media?
Post reply on HN