Live data from Hacker News

Imagen: An AI system that creates photorealistic images from input text

imagen.research.google

201–210 of 233 posts

Re: Imagen: An AI system that creates photorealistic images from input text

#201
post #74

Every single time I see an article about a new AI model that has a section called "societal impact" I know immediately they are not releasing the model, the training set, nothing... It seems to be the kind of bullshit statement that those companies put in place of "we paid $500k training this model and we're not giving it for free to anyone".

Of course that training set was scraped off of everyone else’s web sites, which is fair use. But the model is copyrighted, and I bet you’ll hear an argument that it’s outputs are copyrighted.

The only AI ethics I’m worried about are the lack of anything ethical going on w.r.t intellectual property in AI.

Re: Imagen: An AI system that creates photorealistic images from input text

#202

Earlier quoted context omitted.

Andrew Huberman - How to Cherry Pick Neuroscience "Studies" to Pretend Something is True So You Can Make Lots of Money as a Health Guru Personality

I’ve listened to a number of his podcasts, I’ve never gotten that impression at all. He comes across as a well versed, legit scientist. And there is lots of other evidence to back that up, not just my impression from a few podcasts, publications, tenured professorship at Stanford, etc. But in this case, does it really need extensive scientific study? Just think about it and look around. The thing our entire biologica…

Almost all of his diet/neuroscience/health claims (outside of commonly known things, like "exercise is good for you") are at best a gray area, and at worst just flat wrong. That is because we really don't know much about the brain and its relationship to human experience and behavior - and so anyone claiming to know how to improve your life with "neuroscience" is a con artist. Same with diet, which he also harps on constantly.

For almost every non-obvious claim he makes about something, there exist studies that directly contradict him. And, being a good con artist, he conveniently ignores them. Here are some things he has claimed that are not at all settled science, although he talks as if they are and leaves out the contradictions:

- Artificial sweeteners cause insulin resistance - Light alcohol consumption is detrimental - You can "hack your brain chemicals" with supplements to achieve some desired effect on your mind - We know what role neuromodulators play on really high level cognitive concepts like "creativity" - Increasing the diversity of your gut microbiome has positive health effects

Frankly, no one's life will be any different from listening to Huberman or from doing anything he tells you to do, aside from those things everyone already knows about (like exercising, sleeping an appropriate amount, and not eating a bunch of refined sugar). If there IS some change outside of those, it is just as likely to be the result of random chance than from taking Huberman's advice, because there is no evidence that anything he says (outside of the obvious) is actually beneficial or worthwhile.

To your question - yes...it absolutely DOES need a scientific study. Huberman's claim is porn literally DESTROYS YOUR BRAIN. But again, the studies are terrible and there are many contradicting studies that he would never point out - he's trying to make you think he knows what he's talking about and has something worthwhile to say (when he actually doesn't).

Re: Imagen: An AI system that creates photorealistic images from input text

#203
post #175

Is anyone aware of people doing the same for NSFW images? They can easily wipe out an entire industry. No models to pay and to check for legal ages, infinite possibilities: just write the pic you want, the massive body part you want, how many genitals are involved and boom. You have your image.

[0] Was front page recently. [0] (OBVIOUSLY NSFW) https://news.ycombinator.com/item?id=32572770

This certainly begs the question: When does DALLE2 become viable to generate movie clips of up to 10 minutes in length?

10 min * 60 * 24 FPS = 14400 images ~ 2^13

So maybe 10 years?

Re: Imagen: An AI system that creates photorealistic images from input text

#204
This is amazing but Google/OpenAI haven't released their models and don't seem to plan to. There is stable diffusion which was released and is probably slightly less good but still good if anyone wants to mess with it.

There is a huggingface instance, Collab notebooks, and local running notebooks here. [1] on the stable diffusion subreddit.

Also someone has packaged an exe that runs it with no fuss on computers with Nvidia GPUs that they posted on the media synthesis subreddit[0]

In my limited testing this compares ok to Dalle2. Style shifting works slightly less well and it's hard to force it away from normal images but with a little work it tends to be more accurate to your prompt.

[0] https://grisk.itch.io/stable-diffusion-gui

[1] https://www.reddit.com/r/StableDiffusion/comments/wqaizj/lis...

Re: Imagen: An AI system that creates photorealistic images from input text

#205
post #74

Every single time I see an article about a new AI model that has a section called "societal impact" I know immediately they are not releasing the model, the training set, nothing... It seems to be the kind of bullshit statement that those companies put in place of "we paid $500k training this model and we're not giving it for free to anyone".

Of course that training set was scraped off of everyone else’s web sites, which is fair use. But the model is copyrighted, and I bet you’ll hear an argument that it’s outputs are copyrighted. The only AI ethics I’m worried about are the lack of anything ethical going on w.r.t intellectual property in AI.

Yes, this is how fair use works. Once copyright agrees that you were in the right to make something, you have exclusive ownership over it, it does not get locked open like GPL software does.

Scraping other people's content is already established fair use thanks to Google prevailing against the Author's Guild in front of SCOTUS. I find it difficult to understand how it could be illegal to scrape a bunch of creative works to create a system that generates new works, but legal to scrape a bunch of creative works to let people search two-page excerpts out of them.

Output copyright is very much up in the air. The Copyright Office has rejected copyright registrations claiming the software itself created the work; but presumably this wouldn't apply to a human taking ownership over something they used ML to create. The amount of prompt engineering you have to do to make these kinds of systems alone would count as some kind of creativity. The only real complaint I could see is if the system regurgitated its training data, which would be bog-standard copyright infringement.

Also, related note: I really hate how the whole AI thing is making the FSF sound like the Author's Guild did a decade and change ago. The law is already very clear that you cannot launder an infringement (copying GPL code) through a fair use (trained ML weights). Please do not adopt the arguments of copyright maximalists.

Re: Imagen: An AI system that creates photorealistic images from input text

#206
post #101
post #74

Every single time I see an article about a new AI model that has a section called "societal impact" I know immediately they are not releasing the model, the training set, nothing... It seems to be the kind of bullshit statement that those companies put in place of "we paid $500k training this model and we're not giving it for free to anyone".

It's probably a lot more than that, $500k is like one Google ML engineer's salary.

Right? I feel like novel development of these models looks more like the GPT 10-20M numbers.

Re: Imagen: An AI system that creates photorealistic images from input text

#207
post #121

Earlier quoted context omitted.

Indeed, you can't be trusted not do anything bad with it. Who's going to vet each and every user? Who's going to check every time it's used for the coming 10, 20, 30 years?

If he's racist he can already hire black people to take a photo for an "upcoming action movie" called evil baby. They will be asked to hold guns aiming it at a crib. Then release it on the internet and say he saw 3 black people about to shoot a baby. It would be called out as fake or staged just like an imagen/walle2/openai would be called out as fake. The thing that makes stories real is real people - actual events…

>hire black people to take a photo for an "upcoming action movie" called evil baby. They will be asked to hold guns aiming it at a crib.

As a side note, I'd watch that. I imagine the next scene shows a bloody room and a baby escaping.

Re: Imagen: An AI system that creates photorealistic images from input text

#208
post #183

Earlier quoted context omitted.

There's also a fork that requires a lot less VRAM. I was able to get this working with an Nvidia GTX 1070. https://github.com/basujindal/stable-diffusion You'd clone the fork, then download Stability AI's checkpoint, sd-v1-4.ckpt, from https://huggingface.co/CompVis/stable-diffusion-v-1-4-origin... Follow the instructions in the forked repo, and you should be good to go in a manner of minutes.

The only issue I had was WSL 2 was not using my GPU, so obviously it failed. Couldn't get it working on Windows either, failed at installing dependencies. I'm thinking of dual-booting Linux just to try out stable-diffusion. I normally dev on Linux but don't do GPU work, only my Windows gaming PC has the horsepower to run stable-diffusion and unfortunately I couldn't get it working on Windows. Shame cause I'd love to…

Same. I've spent a number of nights trying variations on the install instructions. Got the windows CUDA drivers (older graphics card), but no matter what I try, conda, or WSL, pytorch refuses to see the CUDA available.

Re: Imagen: An AI system that creates photorealistic images from input text

#209
post #84

Earlier quoted context omitted.

Fortunately we have Stability.AI and they release their image generation model already. Hopefully they'll follow with other projects too. https://stability.ai/blog/stable-diffusion-public-release

There's also a fork that requires a lot less VRAM. I was able to get this working with an Nvidia GTX 1070. https://github.com/basujindal/stable-diffusion You'd clone the fork, then download Stability AI's checkpoint, sd-v1-4.ckpt, from https://huggingface.co/CompVis/stable-diffusion-v-1-4-origin... Follow the instructions in the forked repo, and you should be good to go in a manner of minutes.

Does it make images of lesser quality or does it just take longer for the same quality?

Re: Imagen: An AI system that creates photorealistic images from input text

#210
post #84
post #74

Every single time I see an article about a new AI model that has a section called "societal impact" I know immediately they are not releasing the model, the training set, nothing... It seems to be the kind of bullshit statement that those companies put in place of "we paid $500k training this model and we're not giving it for free to anyone".

Fortunately we have Stability.AI and they release their image generation model already. Hopefully they'll follow with other projects too. https://stability.ai/blog/stable-diffusion-public-release

Much respect to Stability - they followed through on their promise.
Post reply on HN