Live data from Hacker News

Imagen Video: high definition video generation with diffusion models

imagen.research.google

481–490 of 500 posts

Re: Imagen Video: high definition video generation with diffusion models

#481

Earlier quoted context omitted.

I'm sorry you feel so cynical about this. It's absolutely true that Google is profit-seeking, that these models are very expensive to build, and that if there's a competitive advantage to be had, Google should probably try to retain it. But even with that all being true, real people (typically some thoughtful researchers) build these models. And my point is: _there really are ethical reasons to keep large generative…

I'm curious what ethical reasons you think require that new technology only be used in secret and without oversight by trillion dollar companies. This is supposed to be AI safety? "Do whatever you want, just make sure you conceal the results and impede progress and understanding." What reasons necessitate keeping image or video generation models private that wouldn't also argue for keeping animation software or pictu…

> "Do whatever you want, just make sure you conceal the results and impede progress and understanding."

This is not a fair characterization of what's going on here. Google spent a ton of money on researchers & training infra (it's wildly expensive even just hardware-wise) to train these models. It's not different from other proprietary technologies -- they don't owe the public anything here. Providing the research findings + methodology in a paper without the implementation & data is a _tradeoff_ as a participant in the field. If someone else implements the model with their money and uses it for nefarious purposes, that's more acceptable than if they directly use Google's _already known to be flawed_ models.

> I'm curious what ethical reasons you think require that new technology only be used in secret and without oversight by trillion dollar companies. This is supposed to be AI safety?

If I make a chair and I know it's not always safe to sit on, maybe I should not sell that chair. We can talk about this proof-of-concept chair as a research subject, but if you go to build one and use it to prank someone, that's on you.

That's all that's going on here. If the model could be used to generate CSAI, maybe Google doesn't want to be part of that.

> Google is developing image and video generation models and equivalent versions will be open source by the year's end I expect. These models aren't especially dangerous.

Maybe that's the disconnect -- you don't think generative models are dangerous, but they can be, and Google would know because they have entire teams dedicated to AI fairness & safety researching this topic.

It's also not trivial to reproduce these models. Given the cost to simply train even if you had the source data, any organization releasing these models has to have a bit of money and skill. The onus will always be on the team building these models to think about what their ethics are and how they want to proceed knowing there may be negative externalities.

> Yes, people will use them to be racist or mean, same as they use their phones or computers or books or whatever to be those things.

Tools empowering large-scale inauthenticity & disinformation are not comparable to individuals making comments.

Re: Imagen Video: high definition video generation with diffusion models

#482

Earlier quoted context omitted.

It's not as simple as this. Google Search came without Safe Search & other guards at first because _implementing privacy & age controls is hard_. It's a second-order product after the initial product. Bad capabilities (e.g. cyberstalking) are side-effects of a product that "organizes the world's information and makes it universally accessible and useful," and if anything, over time Google has sought build in more saf…

You know safesearch is optional, right? It even disables itself if it knows you're looking for porn. There is nothing that stops children from overriding it. As for learning from the timnit thing I'm pretty sure the only thing people outside Google learned from that is that Google ai "ethicists" all seem to be crazy. Certainly that's the clear vibe on this thread.

> You know safesearch is optional, right? It even disables itself if it knows you're looking for porn. There is nothing that stops children from overriding it.

You can let your kid use Google to look up math lectures without fearing that they would see something slightly traumatizing though, right? That wasn't the case in 1996! The point is that products have varying levels of readiness, and it's totally fair to say "the thing isn't ready, it has too many sharp edges." Especially when the thing could be used at scale.

> As for learning from the timnit thing I'm pretty sure the only thing people outside Google learned from that is that Google ai "ethicists" all seem to be crazy. Certainly that's the clear vibe on this thread.

That's a sad take, but who knows if it's true. HN commenters aren't exactly a representative sample.

Re: Imagen Video: high definition video generation with diffusion models

#483

Earlier quoted context omitted.

Exactly. Probably they already started lobbying against selling high end cheap GPUs to the public. No doubt ether going proof of stake is a huge blow to their agenda. They can't claim all those GPUs are just wasting energy for crypto mining. Now they have to come up with different arguments. I can already see it. Just think of all the energy wasted training AI at home! I can imagine police drones with IR sensors scan…

>advanced AI (same as every other big scientific/engineering achievement) will be predominantly good. They should start teaching the problem of induction in schools, evidently it's needed.

That's a really convoluted way of saying what amounts to "you're dumb".

How about you present a counterargument - why would advanced AI be predominantly bad? Unless, of course, your only counterargument is the classic philosophical statement of "we can't know nuffin".

Re: Imagen Video: high definition video generation with diffusion models

#484

Earlier quoted context omitted.

It'll eventually get to the point where it's high quality and the media you consume will be generated just for you based on your individual preferences, rather than a curated list of already made options made for widespread audiences.

A big part of entertainment's appeal is having an experience/frame of reference to share with other people. Personalized entertainment doesn't offer that. I am also extremely skeptical of the ability/need so serve at individual level instead of niches (as today).

Time will tell.

Re: Imagen Video: high definition video generation with diffusion models

#485

Earlier quoted context omitted.

I'm curious what ethical reasons you think require that new technology only be used in secret and without oversight by trillion dollar companies. This is supposed to be AI safety? "Do whatever you want, just make sure you conceal the results and impede progress and understanding." What reasons necessitate keeping image or video generation models private that wouldn't also argue for keeping animation software or pictu…

> "Do whatever you want, just make sure you conceal the results and impede progress and understanding." This is not a fair characterization of what's going on here. Google spent a ton of money on researchers & training infra (it's wildly expensive even just hardware-wise) to train these models. It's not different from other proprietary technologies -- they don't owe the public anything here. Providing the research fi…

Google uses research, published models, and data that was freely shared with them and iterates on it, making use of their vast budgets and hardware, to develop new models. Then, Google uses those models internally and doesn't share the models. This is a violation of academic norms under the pretense of "safety". As I characterized previously Google is able to do whatever they want, conceal their results, and impede progress and understanding because they aren't sharing their results. You say this isn't a "fair characterization" but it is exactly what is happening - which part is wrong?

You say that Google doesn't "owe the public anything" and that may, or may not, be true from a legal standpoint, but obviously, from a norms, ethical, and moral standpoint Google does have a massive obligation to the public that they are breeching. Google uses the public's data to train, public research, and publicly shared models to iterate on. Then, after building on the shoulders of giants, Google refuses to share what they have built in contravention of the norms that they benefit from.

Regarding your chair metaphor - the "danger" of these models, if there is such, is not that they would hurt the user, like a faulty chair, but that they could be used to hurt others - e.g. a bot army to manipulate public opinion or create fake news. Google isn't building a chair that might break and hurt the user then, but a gun that might hurt others. It's true that guns shouldn't be widely available - not even a die hard libertarian would want a child to have access to a gun, but the entity that sets rules regarding availability is a representative government for the people for whom those rules are being set - not a private company. In other words, if these tools can cause harm they should be regulated by the government, not Google. If the tools are dangerous, that is not an argument that Google should keep them secret.

Re: Imagen Video: high definition video generation with diffusion models

#486

Earlier quoted context omitted.

>advanced AI (same as every other big scientific/engineering achievement) will be predominantly good. They should start teaching the problem of induction in schools, evidently it's needed.

That's a really convoluted way of saying what amounts to "you're dumb". How about you present a counterargument - why would advanced AI be predominantly bad? Unless, of course, your only counterargument is the classic philosophical statement of "we can't know nuffin".

Logical fallacies don't necessarily equal stupidity though sometimes they do.

A bunch of arguments about why AI would be bad have already been advanced. For the economic one, refer to Martin Ford. For the existential one, refer to Bostrom.

Re: Imagen Video: high definition video generation with diffusion models

#487

Earlier quoted context omitted.

> "Do whatever you want, just make sure you conceal the results and impede progress and understanding." This is not a fair characterization of what's going on here. Google spent a ton of money on researchers & training infra (it's wildly expensive even just hardware-wise) to train these models. It's not different from other proprietary technologies -- they don't owe the public anything here. Providing the research fi…

Google uses research, published models, and data that was freely shared with them and iterates on it, making use of their vast budgets and hardware, to develop new models. Then, Google uses those models internally and doesn't share the models. This is a violation of academic norms under the pretense of "safety". As I characterized previously Google is able to do whatever they want, conceal their results, and impede p…

> Google uses research, published models, and data that was freely shared with them and iterates on it, making use of their vast budgets and hardware, to develop new models. Then, Google uses those models internally and doesn't share the models. This is a violation of academic norms under the pretense of "safety".

Google's not doing this (LLM, generative image model) research on academic datasets freely shared with them. They're doing this research on data they gathered at their expense. This is not a violation of academic norms. Again, Google shares a lot of datasets and models, just not LLMs and generative sets trained on problematic source datasets.

> As I characterized previously Google is able to do whatever they want, conceal their results, and impede progress and understanding because they aren't sharing their results. You say this isn't a "fair characterization" but it is exactly what is happening - which part is wrong?

Anyone can do research and not share back to the community. Google _does_ share back to the community in the form of papers (and again, very frequently with models and datasets). If you have the money and expertise to implement the papers, more power to you. Every technology company has some secret sauces they don't share with everyone. That Google may have some of those is not a moral failing.

> from a norms, ethical, and moral standpoint Google does have a massive obligation to the public that they are breeching. Google uses the public's data to train, public research, and publicly shared models to iterate on

From the other end: Google gets user data and has a responsibility to not proliferate that data, no? I wouldn't want them to share a dataset that has my personal data, even if anonymized because there are ways to deanonymize. There are levels to everything, and choosing "I'll release the paper but not the model + data" for some potentially sensitive models seems sane.

> Then, after building on the shoulders of giants, Google refuses to share what they have built in contravention of the norms that they benefit from.

People are building on the shoulders of Google's research all the time, and plenty of companies are doing similar things to Google and being way less open about their work. I mean, every company that trains a big model on data collected from the public -- are they all required to share their models with everyone? Is Cruise sharing their pedestrian detection model? I don't think what you're suggesting could possibly be the standard.

> Regarding your chair metaphor - the "danger" of these models, if there is such, is not that they would hurt the user, like a faulty chair, but that they could be used to hurt others - e.g. a bot army to manipulate public opinion or create fake news.

Sure, I was trying not to be hyperbolic and compare LLMs to guns since they have plenty of awesome use cases (whereas guns really don't). A faulty chair that you set out for anyone to use can hurt people other than the chair's creator / people who are aware of the specific risks. But yeah, seems like you now agree these models have the potential to cause great harm.

> In other words, if these tools can cause harm they should be regulated by the government, not Google

I agree that gov't regulation can be helpful for setting a minimum standard. But I strongly disagree that lack of laws means we should abdicate our own moral responsibilities. If I sell / provide something, I need to be able to sleep at night knowing I didn't make the world worse. Googlers typically try to do this.

Re: Imagen Video: high definition video generation with diffusion models

#488

Earlier quoted context omitted.

Okay Ned Ludd. You're right that technology is only going to get more powerful. But hasn't this always been true? Who decides when it's over the line? The US government? They clearly aren't the best arbiters of judgement, so who gets to decide "sever consequences"?

Yes, as a general rule the government decides when it's ok to forbid you from doing something. If the US government wants, it can make it a crime to train/use these models. Banning the training part is pretty much game over for this industry. I personally hope it happens as soon as possible. Intellectual property theft (without which those models don't exists) shouldn't be allowed.

“Intellectual property” is a propaganda phrase that deliberately confuses copyright, trademark, and patent law in order to create the illusion that ideas can or should have owners. (See https://www.gnu.org/philosophy/not-ipr.en.html) I recommend specifically referring to the concept at hand, in this case copyright.

Personally I disagree, as the models themselves do not contain any material which can be considered a copyright violation, the value of scraping the open web is easily apparent, the ability to prevent scraping wholesale - even given the legal framework to disallow it - seems dubious, and lastly because the collective potential harm caused by restricting just one or a few of the more arguably more ethical nations from this technology pathway is a known unknown, and possibly a very large one at that.

Re: Imagen Video: high definition video generation with diffusion models

#489

Earlier quoted context omitted.

Ironically nearly all ML demos like Stable Diffusion are setup and run for free on Google's Colab, so to claim they don't want to help the field is a little silly.

Stable diffusion can run anywhere

I didn't set it can't, I said the majority of people who have limited skill set with ML and troubleshooting it end up using the Colab version since it's much more straightforward and easy for average people to get started with.

And that's for all ML stuff, Colab has lead to a huge democratization of ML, with notebooks setups for basically any cool demo you see out there.

Re: Imagen Video: high definition video generation with diffusion models

#490

Earlier quoted context omitted.

That's a really convoluted way of saying what amounts to "you're dumb". How about you present a counterargument - why would advanced AI be predominantly bad? Unless, of course, your only counterargument is the classic philosophical statement of "we can't know nuffin".

Logical fallacies don't necessarily equal stupidity though sometimes they do. A bunch of arguments about why AI would be bad have already been advanced. For the economic one, refer to Martin Ford. For the existential one, refer to Bostrom.

Induction is not a logical fallacy, it's a basis for empiricism.
Post reply on HN