Google Imagen 2
101–110 of 194 posts
Re: Google Imagen 2
#102Instead they use
>The robin flew from his swinging spray of ivy on to the top of the wall and he opened his beak and sang a loud, lovely trill, merely to show off. Nothing in the world is quite as adorably lovely as a robin when he shows off - and they are nearly always doing it.
And show off the result being a photograph of a robin, cool. SDXL[0] can do the exact same thing given the same prompt, in fact even SD1.5 would be able to easily[1].
Re: Google Imagen 2
#103I think the competition for text to image services is over and open source, stable diffusion won. It doesn't matter how detailed (or whatever counts as "better") corporate text-to-image products get, stable diffusion is good enough which really is good enough. Unlike the corporate offerings, open source txt2img doesn't have random restrictions (no its not just porn at this point) and actually allows for additional sc…
> Why bother using a product from a company that is notorious for failing to commit to most of their services, when you can run something which produces output that is pretty close (and maybe better) and is free to run and change and train? Because it costs $0.02 per image instead of $1000 on a graphics card and endless buggering around to set up.
Re: Google Imagen 2
#104Wow, Google has really become the IBM of 2005s. All flashy demos, 'call sales' to try anything.
According to Fiona Cicconi, Google’s chief people officer, Google employed 30,000 managers before the recent layoffs. The hard truth is Google needs a Twitter style culling. Take all those billions you're burning and give it to people with a builder mentality, not career sheeple. Unfortunately the same executives who would oversee this are the ones who need to be culled first.
Re: Google Imagen 2
#105they should make it accessible at https://imagen.google like how meta did with https://imagine.meta.com
My kids found it organically and were happily creating all sorts of DALL·E 3 images.
Re: Google Imagen 2
#106I think the competition for text to image services is over and open source, stable diffusion won. It doesn't matter how detailed (or whatever counts as "better") corporate text-to-image products get, stable diffusion is good enough which really is good enough. Unlike the corporate offerings, open source txt2img doesn't have random restrictions (no its not just porn at this point) and actually allows for additional sc…
Why stable diffusion won? Dalle3 and this is miles ahead in understanding scene and put correct text at the right place. This makes the image much more usable without editing.
(DALL-E pretends to do that, but it's actually just using GPT-4 Vision to create a description of the image and then prompting based on that.)
Live editing tools like https://drawfast.tldraw.com/ are increasingly being built on top of Stable Diffusion, and are far and away the most interesting way to interact with image generation models. You can't build that on DALL-E 3.
Re: Google Imagen 2
#107Earlier quoted context omitted.
According to Fiona Cicconi, Google’s chief people officer, Google employed 30,000 managers before the recent layoffs. The hard truth is Google needs a Twitter style culling. Take all those billions you're burning and give it to people with a builder mentality, not career sheeple. Unfortunately the same executives who would oversee this are the ones who need to be culled first.
How did the Twitter-style culling work out for Twitter?
Re: Google Imagen 2
#108I think the competition for text to image services is over and open source, stable diffusion won. It doesn't matter how detailed (or whatever counts as "better") corporate text-to-image products get, stable diffusion is good enough which really is good enough. Unlike the corporate offerings, open source txt2img doesn't have random restrictions (no its not just porn at this point) and actually allows for additional sc…
Still, Stable Diffusion is losing the usability, tooling and integration game. The people who care to make interfaces for it mostly treat it as an expert tool, not something for people who have never heard of image generating AI. Many competing services have better out-of-the-box results (for people who don't know what a negative prompt is), easier hosting, user friendly integrations in tools that matter, better hosted services, etc.
Re: Google Imagen 2
#109This post has more information: https://cloud.google.com/blog/products/ai-machine-learning/i... I can't figure out how to try this thing. The closest I got was this sentence: "To get started with Imagen 2 on Vertex AI, find our documentation or reach out to your Google Cloud account representative to join the Trusted Tester Program."
>> generally available for Vertex AI customers on the allowlist (i.e., approved for access).
Re: Google Imagen 2
#110I think the competition for text to image services is over and open source, stable diffusion won. It doesn't matter how detailed (or whatever counts as "better") corporate text-to-image products get, stable diffusion is good enough which really is good enough. Unlike the corporate offerings, open source txt2img doesn't have random restrictions (no its not just porn at this point) and actually allows for additional sc…
Why stable diffusion won? Dalle3 and this is miles ahead in understanding scene and put correct text at the right place. This makes the image much more usable without editing.
I guess that turns out to be not as important for end users as you'd think.
Anyway, DeepFloyd/IF has great comprehension. It is straightforward to improve that for Stable Diffusion, I cannot tell you exactly why they haven't tried this.