Live data from Hacker News

Imagen Video: high definition video generation with diffusion models

imagen.research.google

301–310 of 500 posts

Re: Imagen Video: high definition video generation with diffusion models

#301

How long until the AI just generates the entire frame buffer on a device? Then you don’t need to design or program anything; the AI just handles all input and output dynamically.

Absolutely believe that's a future we'll get to eventually, no idea on the timeline

Re: Imagen Video: high definition video generation with diffusion models

#302

The concern trolling and gatekeeping about social justice issues coming from the so-called "ethicists" in the AI peanut gallery has been utterly ridiculous. Google claims they don't want to release Imagen because it lacks what can only be called "latent space affirmative action". Stability or someone like it will valiantly release this technology, again and there will be absolutely no harm to anyone. Stop being so to…

There is not ehtical concern. Google will shut it down regardless.

Re: Imagen Video: high definition video generation with diffusion models

#303

Earlier quoted context omitted.

So you would say that if FB fails to censor a video titled "vaccines cause autism" that drives engagement, they are more morally culpable for the content than if Google spends TPU cycles rendering the Imagen prompt input: "detailed video about why vaccines cause autism, scientific, realistic, in the style of a public health announcement" ?

IMO facebook should stop "engagement based" recommendation engines until they have the ability to stop them being abused. Change the platform to show chronological posts from things users have subscribed to. They could perhaps have a curated selection of content that FB employees have screened to be good for general distribution. There is a world of difference between someone manually seeking out and subscribing to a…

That is definitely one opinion that you can hold about Facebook but it's unclear to me how it relates to this post about Google's new generative model

Re: Imagen Video: high definition video generation with diffusion models

#304
post #288

Earlier quoted context omitted.

There is a difference in ease of use. I could never use photoshop to fake something like that even if I wanted to. Further, we have seen harm come from some of this already, there’s a pretty big online community that uses deepfakes to put people in situations they would rather not be in, the most obvious being porn.

> There is a difference in ease of use. I could never use photoshop to fake something like that even if I wanted to. You couldn't, but basically any VFX shop easily could. Point is, it doesn't make anything possible that wasn't already possible, it just makes it more accessible. That's an inevitability with technology, as time goes on. The counter is not to try and suppress it, that has never worked and never will.

What motivation would a VFX shop have to do such a thing, though? (Money, sure, but what's a motive strong enough to be worth commissioning them?)

It's always individuals who want to harm others in this particular way; and individuals don't throw around big-VFX-project amounts of money on petty revenge. But they'd certainly spend $20.

DDoS attacks got a lot (1000x) more commonplace once there were DDoS services that let you buy an hour of attacking someone for $20. Same idea here.

Re: Imagen Video: high definition video generation with diffusion models

#305
post #255

Earlier quoted context omitted.

And copilot-like code, possibly Q1 2023.

"Generate the code base for an advanced diffusion model that can improve on the code base for an advanced diffusion model"

The road to Grey Goo is paved with artificial general intelligence.

Re: Imagen Video: high definition video generation with diffusion models

#306
post #286
post #266

Earlier quoted context omitted.

There is a clear risk from these sorts of models as they get better - I mean recreating specific individuals’ likenesses in compromising images (or even worse, video). We’re not at that point yet, but these things are getting better fast, so it’s only a matter of time. The problem is that there’s no way to mitigate those risks except to keep the model behind an inference-only API, or not release it at all - as soon a…

This is not new. Imagine not releasing tools like curl or nmap because it can be used for hacking. The issue is, as an industry and society, we somehow bought the "safety" and "harm" charade a little bit too much, and somehow think it's a reasonable argument instead of being completely insane.

More like the "safety" and "harm" charade was crammed down our gullets at every turn either by stick or carrot.

Re: Imagen Video: high definition video generation with diffusion models

#307

It's interesting that these models can generate seemingly anything, but the prompt is taken only as a vague suggestion. From the first 15 examples shown to me, only one contained all elements of the prompt, and it was one of the simplest ("an astronaut riding a horse", versus e.g. "a glass ball falling in water" where it's clear it was a water droplet falling and not a glass ball). We're seeing leaps in random capabi…

The way I see it, input being confined to a "text description" is the next immediate problem that needs to be solved. I don't think we can rely on textual inputs for much longer as the human language is too imprecise and/or verbose. It's hard to imagine what exactly the optimal interface would be, but I'm thinking we'll need ways to dictate attributes for each entity being represented, the backdrop, and the view comp…

A "mood board" of inspiration images could be an interesting input method. With deep vision models, we can already separate different "levels" of concepts: a high level subject ("person riding a horse"), textures (the horse hair), medium (painting vs 3d rendering vs photograph), etc. It'd be interesting to have a "smart mood board" that goes from text prompt, to visualizing that hierarchy with different options. Then the user could interactively increase or decrease different parameters, ultimately iterating alongside the computer to realize their creative vision.

Re: Imagen Video: high definition video generation with diffusion models

#308
post #255

Earlier quoted context omitted.

And copilot-like code, possibly Q1 2023.

"Generate the code base for an advanced diffusion model that can improve on the code base for an advanced diffusion model"

Oh no you forgot the important term “but do NOT start turning the universe in to paperclips”.

Re: Imagen Video: high definition video generation with diffusion models

#310
post #282

These videos are not high definition. Stop gaslighting.

High definition is relative. Compared to the previous gen of AI videos, they are extremely crisp.

They look atrociously bad, if some future version of this produces high definition video then i'll be delighted. This seems like a clever but not useful shortcut to 3D. For now this goes in the could but couldn't pile.
Post reply on HN