Live data from Hacker News

Imagen Video: high definition video generation with diffusion models

imagen.research.google

431–440 of 500 posts

Re: Imagen Video: high definition video generation with diffusion models

#431

> However, there are several important safety and ethical challenges remaining. Imagen Video and its frozen T5-XXL text encoder were trained on problematic data. While our internal testing suggest much of explicit and violent content can be filtered out, there still exists social biases and stereotypes which are challenging to detect and filter. We have decided not to release the Imagen Video model or its source code…

Is there any evidence to support the hypothesis that the Russian Federation has been behind the deep fakes? Or is it a case of the good old "common it's obvious, common, commooon"?

Re: Imagen Video: high definition video generation with diffusion models

#432
post #265
post #260

Earlier quoted context omitted.

Sounds like the human brain. Scary!

Scary? Sounds amazing

Amazing if you like the idea of everything your brain is capable of making value of, being made completely obsolete.

Re: Imagen Video: high definition video generation with diffusion models

#433

Can anyone comment on how advanced https://phenaki.video/index.html is? They have an example at the bottom of a 2 minute long video generated from a series of prompts (i.e. a story) which seems more advanced than Google or Meta's recent examples? It didn't get many comments on HN when it was posted.

It seems to be a paper for 2023 conference:

"Under review as a conference paper at ICLR 2023"

So I would say it looks pretty advanced, however they don't use a Diffusion model to generate the images, but an "image conditional video generation", another different approach.

Re: Imagen Video: high definition video generation with diffusion models

#434

Earlier quoted context omitted.

I actually think it's the opposite, AI will probably be writing the stories and humans might occasionally film a few scenes. ~95% of TV shows and movies are cookie-cutter content, with cookie-cutter acting and production values, with the same hooks and the same tropes regurgitated over and over again. Heck they can't even figure out how to make new IP so they keep making reruns of the same old stuff like Star Wars, M…

The last-mile problem applies here too. GPT-3 text is convincing at a distance but when you look closely there is no coherence, no real understanding of plot or emotional dynamics or really anything. TV shows and movies are filled with plot holes and bad writing but it's not that bad. Also I think "a good algorithm" is more than just repetitive content. The plots are reused and generic, but there's real skill involve…

Editors might still have a job :).

Kidding aside, these technologies are amazing, but for a while still they will need a human in the loop selecting, tweaking and editing the output and feeding it back to the contraption for the next iteration.

The question is, for how long?

Re: Imagen Video: high definition video generation with diffusion models

#435

> However, there are several important safety and ethical challenges remaining. Imagen Video and its frozen T5-XXL text encoder were trained on problematic data. While our internal testing suggest much of explicit and violent content can be filtered out, there still exists social biases and stereotypes which are challenging to detect and filter. We have decided not to release the Imagen Video model or its source code…

Someone shot someone with a cheap gun, the cat is out of the bag, gun regulation is pointless, let's let the assault rifles go free is the most american thing I've read all day.

You can't copy paste guns and distribute them for free on the internet. At least not yet.

Re: Imagen Video: high definition video generation with diffusion models

#436
post #171

Earlier quoted context omitted.

> This tool's great grandchild is never going to take a rough idea for a movie and churn out a blockbuster film. What about the tool's nth child though? I think saying it will never do it is a bit much, given what we know about human ingenuity and economic incentives.

I think individual special effects sound very plausible. "Okay, robot, make it so that his arm gets vaporized by an incoming laser, kinda like the same effect in Iron Man 7" is believable to me. But ultimately these things copy other stuff. Artists are often trying to create something that is, at least a bit, new. New is where this approach falls over. By its nature, these things paint from examples. They can design…

To make a blockbuster you don't need to come up with anything new.

Re: Imagen Video: high definition video generation with diffusion models

#437

Earlier quoted context omitted.

They are probably confusing OpenAI with DeepMind, which is owned by Google.

No, I'm very much talking about the Google models. From the original link: "We have taken multiple steps to minimize these concerns, for example in internal trials, we apply input text prompt filtering, and output video content filtering. However, there are several important safety and ethical challenges remaining. Imagen Video and its frozen T5-XXL text encoder were trained on problematic data. While our internal te…

[deleted]

Re: Imagen Video: high definition video generation with diffusion models

#438

I agree with many of the arguments in this thread: that model-gatekeeping while publishing approaches seems insincere and just seems like it's daring bad actors to replicate. However, a common refrain is that AI is like tools like hammers or knives and can be used for good or misused for evil. The potential for weaponizing AI is much much more so than a hammer or a knife. And it's greater than 3D-printing (of guns),…

Yeah man, great take. Should we drop a nuke on a city or open source DALL-E ? Seems about equally destructive.

Re: Imagen Video: high definition video generation with diffusion models

#439

> However, there are several important safety and ethical challenges remaining. Imagen Video and its frozen T5-XXL text encoder were trained on problematic data. While our internal testing suggest much of explicit and violent content can be filtered out, there still exists social biases and stereotypes which are challenging to detect and filter. We have decided not to release the Imagen Video model or its source code…

> [...] there still exists social biases and stereotypes which are challenging to detect and filter.

If Google filters them, wouldn't the result be still biased and stereotyped, just along Google's biases? "I reject your biases and substitute my own!"

Re: Imagen Video: high definition video generation with diffusion models

#440
post #318

Earlier quoted context omitted.

This thread is cathartic. I've been feeling uncomfortable with the level of control being sought over the usage of these tools for a while, but didn't want to ruffle the wrong feathers while just getting into AI as a hobby. I think there will be a pretty short window in which all of this hand-waving will be taken seriously. Not because AI won't be used for terrible things (I'm sure it already is) but because consumer…

Exactly. Probably they already started lobbying against selling high end cheap GPUs to the public. No doubt ether going proof of stake is a huge blow to their agenda. They can't claim all those GPUs are just wasting energy for crypto mining. Now they have to come up with different arguments. I can already see it. Just think of all the energy wasted training AI at home! I can imagine police drones with IR sensors scan…

>advanced AI (same as every other big scientific/engineering achievement) will be predominantly good.

They should start teaching the problem of induction in schools, evidently it's needed.

Post reply on HN