Edit: Please ignore. They hadn't rolled the new model out to my account yet. The announcement blog post is a bit misleading saying you can try it today. -- Comparison with Leonardo.Ai. ChatGPT: https://chatgpt.com/share/67e2fb21-a06c-8008-b297-07681dddee... ChatGPT again (direct one shot): https://chatgpt.com/share/67e2fc44-ecc8-8008-a40f-e1368d306e... ChatGPT again (using word "photorealistic instead of "photo"): ht…
The ChatGPT examples don't look like the new Image Gen model yet. The text on the dog collar isn't very good.
4o Image Generation
61–70 of 629 posts
Re: 4o Image Generation
#62Is it live yet? Have been trying it out and am still getting poor results on text generation.
Might take a day or two before it's available in general.
Re: 4o Image Generation
#63OpenAI's livestream of GPT-4o Image Generation shows that it is slowwwwwwwwww (maybe 30 seconds per image, which Sam Altman had to spin "it's slow but the generated images are worth it"). Instead of using a diffusion approach, it appears to be generating the image tokens and decoding them akin to the original DALL-E ( https://openai.com/index/dall-e/ ), which allows for streaming partial generations from top to botto…
LLMs are autoregressive, so they can't be (multi-modality) integrated with diffusion image models, only with autoregressive image models (which generate an image via image tokens). Historically those had lower image fidelity than diffusion models. OpenAI now seems to have solved this problem somehow. More than that, they appear far ahead of any available diffusion model, including Midjourney and Imagen 3. Gemini "int…
Re: 4o Image Generation
#64Earlier quoted context omitted.
The ChatGPT examples don't look like the new Image Gen model yet. The text on the dog collar isn't very good.
Apparently it rolls out today to Plus (which I have). I followed the "Try in ChatGPT" link at the top of the post
Re: 4o Image Generation
#65> Introducing 4o Image Generation: [...] our most advanced image generator yet Then google: > Gemini 2.5: Our most intelligent AI model > Introducing Gemini 2.0 | Our most capable AI model yet I could go on forever. I hope this trend dies and apple starts using something effective so all the other companies can start copying a new lexicon.
Which is especially relevant when it's not obvious which product is the latest and best just looking at the names. Lots of tech naming fails this test from Xbox (Series X vs S) to OpenAI model names (4o vs o1-pro).
Here they claim 4o is their most capable image generator which is useful info. Especially when multiple models in their dropdown list will generate images for you.
Re: 4o Image Generation
#66Is it live yet? Have been trying it out and am still getting poor results on text generation.
It seems like an odd way to name/announce it, there's nothing obvious to distinguish it from what was already there (i.e. 4o making images) so I have no idea if there is a UI change to look for, or just keep trying stuff until it seems better?
Re: 4o Image Generation
#67Tried it, the "compise armporressed" and "Pros: made bord reqotons" didn't impress me in the slightest.
Re: 4o Image Generation
#68Re: 4o Image Generation
#69Earlier quoted context omitted.
The ChatGPT examples don't look like the new Image Gen model yet. The text on the dog collar isn't very good.
Apparently it rolls out today to Plus (which I have). I followed the "Try in ChatGPT" link at the top of the post
Re: 4o Image Generation
#70> Introducing 4o Image Generation: [...] our most advanced image generator yet Then google: > Gemini 2.5: Our most intelligent AI model > Introducing Gemini 2.0 | Our most capable AI model yet I could go on forever. I hope this trend dies and apple starts using something effective so all the other companies can start copying a new lexicon.