Earlier quoted context omitted.
That's the point. With the old models they all failed to produce a wine glass that is completley to the brim full. Because you can't find that a lot in the data they used for training.
Imagine if they just actually trained the model on a bunch of photographs of a full glass of wine, knowing of this litmus test
4o Image Generation
191–200 of 629 posts
Re: 4o Image Generation
#192...Once the wait time is up, I can generate the corrected version with exactly eight characters: five mice, one elephant, one polar bear, and one giraffe in a green turtleneck. Let me know if you'd like me to try again later!
Re: 4o Image Generation
#193Earlier quoted context omitted.
Agreed. It seems totally unnatural that a couple of nerds high-five awkwardly.
Not awkward. Anatomically uncanny and physically impossible.
Re: 4o Image Generation
#194Earlier quoted context omitted.
It still can't generate a full glass of wine. Even in follow up questions it failed to manipulate the image correctly.
https://i.imgur.com/xsFKqsI.png "Draw a picture of a full glass of wine, ie a wine glass which is full to the brim with red wine and almost at the point of spilling over... Zoom out to show the full wine glass, and add a caption to the top which says "HELL YEAH". Keep the wine level of the glass exactly the same."
Re: 4o Image Generation
#195To avoid confusion, why not always use a general AI model upfront, then depending on the user's prompt, redirect it to a specific model?
As to why they don't automatically detect when reasoning could be appropriate and then switch to o3, I don't know, but I'd assume it's about cost (and for most users the output quality is negligible). 4o can do everything, it's just not great at "logic".
Re: 4o Image Generation
#196Earlier quoted context omitted.
You know the images themselves don’t get shared in links like that, right? (It even tells you so when you make the link.)
I created a shared link just now, was not presented with any such warning, and have the same problem with the image not showing up: https://chatgpt.com/share/67e319dd-bd08-8013-8f9b-6f5140137f...
Re: 4o Image Generation
#197Is there any way to see whether a given prompt was serviced by 4o or Dall-E? Currently, my prompts seem to be going to the latter still, based on e.g. my source image being very obviously looped through a verbal image description and back to an image, compared to gemini-2.0-flash-exp-image-generation. A friend with a Plus plan has been getting responses from either. The long-term plan seems to be to move to 4o comple…
Re: 4o Image Generation
#198Re: 4o Image Generation
#199What's important about this new type of image generation that's happening with tokens rather than with diffusion, is that this is effectively reasoning in pixel space. Example: Ask it to draw a notepad with an empty tic-tac-toe, then tell it to make the first move, then you make a move, and so on. You can also do very impressive information-conserving translations, such as changing the drawing style, but also stuff l…
Re: 4o Image Generation
#200Earlier quoted context omitted.
Can you do this with the prompt of a cow jumping over the moon? I can’t ever seem to get it to make the cow appear to be above the moon. Always literally covering it or to the side etc.
https://chatgpt.com/share/67e31a31-3d44-8011-994e-b7f8af7694... got it on the second try.