Live data from Hacker News

Imagen Video: high definition video generation with diffusion models

imagen.research.google

281–290 of 500 posts

Re: Imagen Video: high definition video generation with diffusion models

#283

The concern trolling and gatekeeping about social justice issues coming from the so-called "ethicists" in the AI peanut gallery has been utterly ridiculous. Google claims they don't want to release Imagen because it lacks what can only be called "latent space affirmative action". Stability or someone like it will valiantly release this technology, again and there will be absolutely no harm to anyone. Stop being so to…

Wrongfully or not, people blame Facebook for inflammatory content posted by humans on their platform. How much worse would it be if that content was generated with FB-trained models?

Re: Imagen Video: high definition video generation with diffusion models

#284

The concern trolling and gatekeeping about social justice issues coming from the so-called "ethicists" in the AI peanut gallery has been utterly ridiculous. Google claims they don't want to release Imagen because it lacks what can only be called "latent space affirmative action". Stability or someone like it will valiantly release this technology, again and there will be absolutely no harm to anyone. Stop being so to…

I agree basically completely, but there’s now a cottage industry of AI Ethics professionals whose real job is to provide a smoke screen for the “cake and eat it too” that the big shops want on this kit: peer review and open source contributions and an academic atmosphere when it suits them, proprietary when it doesn’t. Those folks are a lobby now. The thing about owning the data sets and the huge TPU/A100 clusters is…

Just because there are professionals doesn't mean we have to respect their arguments. There are people who get paid to be antivaxxers, doesn't mean we have to listen to them.

"What have you done this week?"

Re: Imagen Video: high definition video generation with diffusion models

#285
post #266

The concern trolling and gatekeeping about social justice issues coming from the so-called "ethicists" in the AI peanut gallery has been utterly ridiculous. Google claims they don't want to release Imagen because it lacks what can only be called "latent space affirmative action". Stability or someone like it will valiantly release this technology, again and there will be absolutely no harm to anyone. Stop being so to…

There is a clear risk from these sorts of models as they get better - I mean recreating specific individuals’ likenesses in compromising images (or even worse, video). We’re not at that point yet, but these things are getting better fast, so it’s only a matter of time. The problem is that there’s no way to mitigate those risks except to keep the model behind an inference-only API, or not release it at all - as soon a…

> There is a clear risk from these sorts of models as they get better - I mean recreating specific individuals’ likenesses in compromising images (or even worse, video).

This has been possible without AI for a very very long time now (just open photoshop, etc). It barely ever happens, and society hasn't collapsed.

I keep seeing this argument come up and it baffles me that informed technologists take it seriously, as if it were impossible to convincingly manipulate images before DALL-E came around.

Re: Imagen Video: high definition video generation with diffusion models

#286
post #266

The concern trolling and gatekeeping about social justice issues coming from the so-called "ethicists" in the AI peanut gallery has been utterly ridiculous. Google claims they don't want to release Imagen because it lacks what can only be called "latent space affirmative action". Stability or someone like it will valiantly release this technology, again and there will be absolutely no harm to anyone. Stop being so to…

There is a clear risk from these sorts of models as they get better - I mean recreating specific individuals’ likenesses in compromising images (or even worse, video). We’re not at that point yet, but these things are getting better fast, so it’s only a matter of time. The problem is that there’s no way to mitigate those risks except to keep the model behind an inference-only API, or not release it at all - as soon a…

This is not new. Imagine not releasing tools like curl or nmap because it can be used for hacking.

The issue is, as an industry and society, we somehow bought the "safety" and "harm" charade a little bit too much, and somehow think it's a reasonable argument instead of being completely insane.

Re: Imagen Video: high definition video generation with diffusion models

#287

The concern trolling and gatekeeping about social justice issues coming from the so-called "ethicists" in the AI peanut gallery has been utterly ridiculous. Google claims they don't want to release Imagen because it lacks what can only be called "latent space affirmative action". Stability or someone like it will valiantly release this technology, again and there will be absolutely no harm to anyone. Stop being so to…

Wrongfully or not, people blame Facebook for inflammatory content posted by humans on their platform. How much worse would it be if that content was generated with FB-trained models?

People blame Facebook for their intentional, and continued choice to use algorithms that amplify and surface inflammatory content.

Re: Imagen Video: high definition video generation with diffusion models

#288
post #266

Earlier quoted context omitted.

There is a clear risk from these sorts of models as they get better - I mean recreating specific individuals’ likenesses in compromising images (or even worse, video). We’re not at that point yet, but these things are getting better fast, so it’s only a matter of time. The problem is that there’s no way to mitigate those risks except to keep the model behind an inference-only API, or not release it at all - as soon a…

> There is a clear risk from these sorts of models as they get better - I mean recreating specific individuals’ likenesses in compromising images (or even worse, video). This has been possible without AI for a very very long time now (just open photoshop, etc). It barely ever happens, and society hasn't collapsed. I keep seeing this argument come up and it baffles me that informed technologists take it seriously, as…

There is a difference in ease of use. I could never use photoshop to fake something like that even if I wanted to.

Further, we have seen harm come from some of this already, there’s a pretty big online community that uses deepfakes to put people in situations they would rather not be in, the most obvious being porn.

Re: Imagen Video: high definition video generation with diffusion models

#289

Earlier quoted context omitted.

Imagine you’re watching a show, it’s really funny and you’re enjoying it. You’re streaming it, but you’d probably have paid a few dollars to rent it back in the Blockbuster days. You’re then told that the show was produced by an AI. Do you suddenly lose interest because you don’t want to watch something produced by an AI? Or is your hypothesis that an AI could never produce a show that you liked to that degree? If yo…

You may want to familiarize yourself with this thought experiment and think how a slightly modified version applies to AIs and their output: https://en.wikipedia.org/wiki/Experience_machine As to whether I am an outlier: Hundreds of thousands of people worldwide watch Magnus Carlsen. How many have watched AlphaZero play chess when it came about and how many watch it when it ceased to be a novelty?

Totally different. Watching a display of skill, where you marvel at how much better the demonstrator is than yourself obviously has no value if the demonstrator is a machine, but then it is plainly visible that the activity has little intrinsic entertainment value and entertainment value comes from the story and personal arc of the performer. This is different from a movie where nobody really cares about the personal arc of the actor, and people are completely happy to watch an animated film where there isn't even a real actor on display.

Re: Imagen Video: high definition video generation with diffusion models

#290
post #267

How long until the AI just generates the entire frame buffer on a device? Then you don’t need to design or program anything; the AI just handles all input and output dynamically.

Imagine you click a youtube video in a bad network envoirment, then the server sends like an alt tag equivalent for the video as a promnt, and the Neural Engine chip inside your phone create the first seconds of the video while it loads. We're fay away from it now, but I've seen less sketchy solutions being implemented.

According to this source, step 3 of the cascading model generates a 16 frame video at 24×48 resolution. So instead of sending a text prompt YouTube could almost just as easily send 16 downsampled frames of the beginning of the video that your Neural Engine chip could work on instead.
Post reply on HN