Live data from Hacker News

Imagen Video: high definition video generation with diffusion models

imagen.research.google

291–300 of 500 posts

Re: Imagen Video: high definition video generation with diffusion models

#291

The concern trolling and gatekeeping about social justice issues coming from the so-called "ethicists" in the AI peanut gallery has been utterly ridiculous. Google claims they don't want to release Imagen because it lacks what can only be called "latent space affirmative action". Stability or someone like it will valiantly release this technology, again and there will be absolutely no harm to anyone. Stop being so to…

Wrongfully or not, people blame Facebook for inflammatory content posted by humans on their platform. How much worse would it be if that content was generated with FB-trained models?

Facebook is a recommendation engine. The problem isn't so much the content, its that facebook chooses to show it to people who did not actively seek it out. You never see people complain at Chrome for showing the content or nginx for hosting it.

Recommendation engines are more responsible than basic infrastructure.

Re: Imagen Video: high definition video generation with diffusion models

#292
post #287

Earlier quoted context omitted.

Wrongfully or not, people blame Facebook for inflammatory content posted by humans on their platform. How much worse would it be if that content was generated with FB-trained models?

People blame Facebook for their intentional, and continued choice to use algorithms that amplify and surface inflammatory content.

I think it's a hard problem and I'm not sure what the right solution is; clearly extremism is a problem but I can't say I'm 100% happy with Facebook being the final judge of Truth.

Regardless, though, it is unambiguous that FB's role in "making" problematic UGC is much less direct than Google's role in making Imagen outputs.

Re: Imagen Video: high definition video generation with diffusion models

#294

Earlier quoted context omitted.

The way I see it, input being confined to a "text description" is the next immediate problem that needs to be solved. I don't think we can rely on textual inputs for much longer as the human language is too imprecise and/or verbose. It's hard to imagine what exactly the optimal interface would be, but I'm thinking we'll need ways to dictate attributes for each entity being represented, the backdrop, and the view comp…

Like an AI-assisted photoshop, but that's not restricted by language, only interactivity. Down the line you'll need a direct mind meld because some ideas don't have words to describe them, but that's not the "next immediate" problem.

People quickly created photoshop plugins to integrate stable diffusion when SD came out

https://twitter.com/wbuchw/status/1563162131024920576

Re: Imagen Video: high definition video generation with diffusion models

#295

Earlier quoted context omitted.

Wrongfully or not, people blame Facebook for inflammatory content posted by humans on their platform. How much worse would it be if that content was generated with FB-trained models?

Facebook is a recommendation engine. The problem isn't so much the content, its that facebook chooses to show it to people who did not actively seek it out. You never see people complain at Chrome for showing the content or nginx for hosting it. Recommendation engines are more responsible than basic infrastructure.

So you would say that if FB fails to censor a video titled "vaccines cause autism" that drives engagement, they are more morally culpable for the content than if Google spends TPU cycles rendering the Imagen prompt input: "detailed video about why vaccines cause autism, scientific, realistic, in the style of a public health announcement"

?

Re: Imagen Video: high definition video generation with diffusion models

#296
post #266

The concern trolling and gatekeeping about social justice issues coming from the so-called "ethicists" in the AI peanut gallery has been utterly ridiculous. Google claims they don't want to release Imagen because it lacks what can only be called "latent space affirmative action". Stability or someone like it will valiantly release this technology, again and there will be absolutely no harm to anyone. Stop being so to…

There is a clear risk from these sorts of models as they get better - I mean recreating specific individuals’ likenesses in compromising images (or even worse, video). We’re not at that point yet, but these things are getting better fast, so it’s only a matter of time. The problem is that there’s no way to mitigate those risks except to keep the model behind an inference-only API, or not release it at all - as soon a…

>There is a clear risk from these sorts of models as they get better - I mean recreating specific individuals’ likenesses in compromising images

And the risk behind that is...?

If you drill down with such claims the core is always "someone might use this to lie online" and the proposed solution every single time is: more surveillance. End anonymity. Have a Facebook account required to use the internet. Real name and real face policies for every online interaction.

Re: Imagen Video: high definition video generation with diffusion models

#297

Earlier quoted context omitted.

Facebook is a recommendation engine. The problem isn't so much the content, its that facebook chooses to show it to people who did not actively seek it out. You never see people complain at Chrome for showing the content or nginx for hosting it. Recommendation engines are more responsible than basic infrastructure.

So you would say that if FB fails to censor a video titled "vaccines cause autism" that drives engagement, they are more morally culpable for the content than if Google spends TPU cycles rendering the Imagen prompt input: "detailed video about why vaccines cause autism, scientific, realistic, in the style of a public health announcement" ?

IMO facebook should stop "engagement based" recommendation engines until they have the ability to stop them being abused. Change the platform to show chronological posts from things users have subscribed to. They could perhaps have a curated selection of content that FB employees have screened to be good for general distribution.

There is a world of difference between someone manually seeking out and subscribing to a misinformation source than FB automatically suggesting it to them.

Re: Imagen Video: high definition video generation with diffusion models

#298
post #267

Earlier quoted context omitted.

Imagine you click a youtube video in a bad network envoirment, then the server sends like an alt tag equivalent for the video as a promnt, and the Neural Engine chip inside your phone create the first seconds of the video while it loads. We're fay away from it now, but I've seen less sketchy solutions being implemented.

I wonder if this is an area actively being researched, using models like these for video compression?

There was a very cool project recently using StableDiffusion to compress images better than JPEG.

Also, there's some interesting work with ML taking diffused light from around a corner and recovering the original pre-diffused silhouette.

In many ways, this is how we've learned the visual cortex is working.

The amount of actual neutral data you are seeing is way less than you'd think given your perceived visual fidelity.

The only practical issue is that distribution of AI hardware in consumer devices is going to noticeably lag behind POC on compounding cutting edge hardware in research environments, and no one wants to invest into obsolescence.

Maybe it will happen in the cellphone market though given the hardware refresh rates from carrier subsidies.

Re: Imagen Video: high definition video generation with diffusion models

#299
post #47

And there you have it. As an aspiring filmmaker and an AI researcher, I'm going to relish the next decade or so where my talents are still relevant. We're entering the golden age of art, where the AIs are just good enough to be used as tools to create more and more creative things, but not good enough yet to fully replace the artist. I'm excited for the golden age, and uncertain about what comes after it's over, but…

> fully replace the artist I doubt the artist would ever be "fully" replaced, or even mostly replaced. People very much care about the artist when they buy art in pretty much any form. Mass produced art has always been a thing, but I'm not alone in not wanting some $15 print from IKEA on my wall, even if it were to be unique and beautiful. Etsy successfully sells tons of hand-made goods, even though factories can pro…

Thanks for validating my hatred of those IKEA paintings lol. Close-up zebras, black and white picture of Amsterdam with a red bicycle...

Re: Imagen Video: high definition video generation with diffusion models

#300
post #288

Earlier quoted context omitted.

> There is a clear risk from these sorts of models as they get better - I mean recreating specific individuals’ likenesses in compromising images (or even worse, video). This has been possible without AI for a very very long time now (just open photoshop, etc). It barely ever happens, and society hasn't collapsed. I keep seeing this argument come up and it baffles me that informed technologists take it seriously, as…

There is a difference in ease of use. I could never use photoshop to fake something like that even if I wanted to. Further, we have seen harm come from some of this already, there’s a pretty big online community that uses deepfakes to put people in situations they would rather not be in, the most obvious being porn.

> There is a difference in ease of use. I could never use photoshop to fake something like that even if I wanted to.

You couldn't, but basically any VFX shop easily could. Point is, it doesn't make anything possible that wasn't already possible, it just makes it more accessible. That's an inevitability with technology, as time goes on. The counter is not to try and suppress it, that has never worked and never will.

Post reply on HN