Live data from Hacker News

Google Imagen 2

cloud.google.com

21–30 of 194 posts

Re: Google Imagen 2

#21
post #17
post #11

Earlier quoted context omitted.

Stability AI has gaps in SDXL for text, but they seem to do a better job with Deep Floyd ( https://github.com/deep-floyd/IF ). I have done a lot of interesting text things with Deep Floyd

Looks good. But 24GB of vram is quite a lot for 1024x1024

This is a pixel diffusion model that doesn't use latent space encoding, hence the memory requirements. Besides, good prompt understanding requires large transformers for text encoding, usually far larger than the image generation part. DF IF is using T5.

You can use Harrlogos XL to produce text with SDXL, although it's mostly limited to short captions and logos. The other way (controlnets) is more involved. (and is actually useful)

Re: Google Imagen 2

#25
post #20

Wow, Google has really become the IBM of 2005s. All flashy demos, 'call sales' to try anything.

According to Fiona Cicconi, Google’s chief people officer, Google employed 30,000 managers before the recent layoffs. The hard truth is Google needs a Twitter style culling. Take all those billions you're burning and give it to people with a builder mentality, not career sheeple. Unfortunately the same executives who would oversee this are the ones who need to be culled first.

Re: Google Imagen 2

#26

Google desperately needs to get their platform/docs in order. It is incredibly difficult to use any of their new AI stuff. I have access to Imagen (which was a rodeo to get on its own), but do not know if it v1 or v2 for example.

They need to ditch Sundar, I don't know what the hell they are thinking. Google so badly needs reorganization.

Re: Google Imagen 2

#27
post #9
post #3

This would have been an epic release two years ago, but there are now many well-established models in this area (DALL-E, Midjourney, Stable Diffusion). It would be great to see some comparisons or benchmarks to show Imagen 2 is a better alternative. As it stands, it's hard for me to tell if this is worth switching to.

> it's hard for me to tell I can only compare it to Stable Diffusion. But Imagen2 seems significant more advanced. Try to do anything with text and SDxl. It's not easy and often messes up. I don't think you can get a clean logo with multiple text areas on sdxl. Look at the prompt and image of the robin. That is mighty impressive.

> I can only compare it to Stable Diffusion. But Imagen2 seems significant more advanced.

I wouldn't say this until we are able to try it for ourselves. As we know, Google is prone to severe cherry picking and deceptive marketing.

Re: Google Imagen 2

#29
For the peer comments

- https://cloud.google.com/vertex-ai (marketing page)

- https://cloud.google.com/vertex-ai/docs (docs entry point)

- https://console.cloud.google.com/vertex-ai (cloud console)

- https://console.cloud.google.com/vertex-ai/model-garden (all the models)

- https://console.cloud.google.com/vertex-ai/generative (studio / playground)

VertexAI is the umbrella for all of the Google models available through their cloud platform.

It still seems there is confusion (at google) about this being TTP or GA. Docs say both, the studio has a request access link.

more... this page has a table with features and current access levels: https://cloud.google.com/vertex-ai/docs/generative-ai/image/...

Seems that some features are GA while others are still in early access, in particular image generation is still EA, or what they call "Restricted GA"

Re: Google Imagen 2

#30

But how do we use it? Yet another documentation release by googling, promising impressive things that we cannot actually use, while the competition is readily available.

I still cannot believe they missed one of the most critical parts of this release - clear and simple instructions on how to use it. How do they even hope to get adoption without that is unclear to me.
Post reply on HN