Live data from Hacker News

Imagen Video: high definition video generation with diffusion models

imagen.research.google

81–90 of 500 posts

Re: Imagen Video: high definition video generation with diffusion models

#82
post #78

Earlier quoted context omitted.

Think about where this stuff was 2 years ago and then think about where it will be 2 years from now.

Relationships between objects has been a problem with computer vision for a long time. 10 years ago: https://karpathy.github.io/2012/10/22/state-of-computer-visi... Now: https://arxiv.org/pdf/2204.13807 Given that this is what makes photos and videos interesting I think it's still a while before artists are automated.

Take a look at Flamingo "solving" the joke: https://pbs.twimg.com/media/FSFwYL7WUAEgxqQ?format=jpg&name=...

Re: Imagen Video: high definition video generation with diffusion models

#84

I’m going to post an Ask HN about what am I supposed to do when I’m “disrupted”. I work in film / video / CG where the bread and butter is short form advertising for Youtube, Instagram and TV. It’s painfully obvious that in 1 year the job might be exceedingly more difficult than it is now.

There's a huge gap between "that's pretty cool" and a feature length film. People want to create specific stories with specific scenes in specific places that look a specific way. A "Couple kissing in the rain " prompt isn't going to produce something people are going to pay to see.

It's more likely that you're still going to be filming/editing/animating but will have an AI layer on top that produces extra effects or generates pieces of a scene. Think "green screen plus", vs fully AI entertainment.

People will over-hype this tech like they did with voice and driverless cars but don't let it scare you. Everything is possible, but it's like a person from the 1920's telling everyone the internet will be a thing. Yes it's correct, but also irrelevant at the same time. You already have AI assisted software being used in your industry. Just expect more of that and learn how to use the tools.

Re: Imagen Video: high definition video generation with diffusion models

#86

"We have decided not to release the Imagen Video model or its source code until these concerns are mitigated" Okay then why even post it in the first place? What exactly is Google going to do with this model?

It's a research activity.

Google and Meta and Microsoft all have research teams working on AI.

Putting out papers like this helps keep their existing employees happy (since they get to take credit for their work) and helps attract other skilled employees as well.

Re: Imagen Video: high definition video generation with diffusion models

#87

Earlier quoted context omitted.

Top two comments are creatives wondering about their future jobs. Ai ethicists have brought up concerns regarding intentional misuse like misinformation. The technology is super cool. Cat is out of the bag. Just like we couldn't really make cryptography illegal, this stuff shouldn't be either. But I dislike how everyone is pretending that AI ethicists and others are completely unfounded just because it is popular to…

It’s impressive that the small videos are generated this way but the videos themselves are obviously ML generated as they are distorted, a lot like the other art, you can kinda tell it’s the computer. I’m not seeing the ethical issues. I mean cameras disrupted lots of jobs. In general that’s what all technology does everyday. What’s different about this technology?

The difference with this technology are the unlimited possibilities to generate any type of video content with low knowledge barrier and relatively low investment required. The ethical issue is not about how this technology could disrupt the video job market, but how powerful content it can create literally on the fly. I mean, you can tell it's computer generated ... for now.

Re: Imagen Video: high definition video generation with diffusion models

#88
post #66

I’m going to post an Ask HN about what am I supposed to do when I’m “disrupted”. I work in film / video / CG where the bread and butter is short form advertising for Youtube, Instagram and TV. It’s painfully obvious that in 1 year the job might be exceedingly more difficult than it is now.

Start making content and charging for it. You no longer need institutional capital to make a Disney- or Pixar-like experience. Small creators will win under this new regime of tools. It's a democratizing force.

> It's a democratizing force.

I'm wondering why the open source community doesn't get this. So many voices were raised against Codex. Now artists against Diffusion models. But the model itself is a distillation of everything we created, it can compactly encode it and recreate it in any shape and form we desire. That means everyone gets to benefit, all skills are available for everyone, all tailored to our needs.

Re: Imagen Video: high definition video generation with diffusion models

#89

Earlier quoted context omitted.

Audio is the next thing that Stability AI is dropping, then video. In a few months you'll be able to conjure up anything you want if you have a few GPU cores. Pretty incredible.

I won’t be impressed until it can generate smells.

You joke, but that is in the works as well (would require special hardware though) https://ai.googleblog.com/2022/09/digitizing-smell-using-mol...
Post reply on HN