Earlier quoted context omitted.
Can't replicate. Maybe the rollout is staggered? Using Plus from Europe, it's consistently giving me a half full glass.
You might still be on DALL-E. My account is if you use ChatGPT. I switched over to the sora.com domain and now I have access to it.
4o Image Generation
341–350 of 629 posts
Re: 4o Image Generation
#342Earlier quoted context omitted.
Hmmm, I wanted to do that tic tac toe example, and it failed to create a 3x3 grid, instead creating a 5x5 (?) grid with two first moves marked. https://chatgpt.com/share/67e32d47-eac0-8011-9118-51b81756ec...
Your images say "Created with DALL-E", so you have not tried out the new model yet. I think they are gradually rolling it out.
Re: 4o Image Generation
#343OpenAI's livestream of GPT-4o Image Generation shows that it is slowwwwwwwwww (maybe 30 seconds per image, which Sam Altman had to spin "it's slow but the generated images are worth it"). Instead of using a diffusion approach, it appears to be generating the image tokens and decoding them akin to the original DALL-E ( https://openai.com/index/dall-e/ ), which allows for streaming partial generations from top to botto…
Re: 4o Image Generation
#344Is there any way to see whether a given prompt was serviced by 4o or Dall-E? Currently, my prompts seem to be going to the latter still, based on e.g. my source image being very obviously looped through a verbal image description and back to an image, compared to gemini-2.0-flash-exp-image-generation. A friend with a Plus plan has been getting responses from either. The long-term plan seems to be to move to 4o comple…
the native just.. works
Re: 4o Image Generation
#345Re: 4o Image Generation
#346What's important about this new type of image generation that's happening with tokens rather than with diffusion, is that this is effectively reasoning in pixel space. Example: Ask it to draw a notepad with an empty tic-tac-toe, then tell it to make the first move, then you make a move, and so on. You can also do very impressive information-conserving translations, such as changing the drawing style, but also stuff l…
> truly generative UI, where the model produces the next frame of the app I built this exact thing last month, demo: https://universal.oroborus.org (not viable on phone for this demo, fine on tablet or computer) Also see discussion and code at: http://github.com/snickell/universal I wasn't really planning to share/release it today, but, heck, why not. I started with bitmap-style generative image models, but because t…
Re: 4o Image Generation
#347Earlier quoted context omitted.
> truly generative UI, where the model produces the next frame of the app I built this exact thing last month, demo: https://universal.oroborus.org (not viable on phone for this demo, fine on tablet or computer) Also see discussion and code at: http://github.com/snickell/universal I wasn't really planning to share/release it today, but, heck, why not. I started with bitmap-style generative image models, but because t…
Do you have any demo videos?
You can watch "sped up" past sessions by other people who used this demo here, which is kind of like a demo video: https://universal.oroborus.org/gallery
But the gallery feature isn't really there today, it shows all the "one-click and bounce sessions", and its hard to find signal in the noise.
I'll probably submit a "Show HN" when I have the gallery more together, and I think its a great idea to pick a multi-click gallery sequence and upload it as a video.
Re: 4o Image Generation
#348Earlier quoted context omitted.
I hate modern marketing trends. This one isn't even my biggest gripe. If I could eliminate any word from the English language forever, it would be "effortlessly".
Idk, right now I think I'd eliminate "blazingly fast" from software engineering vocabulary.
Re: 4o Image Generation
#349https://news.ycombinator.com/item?id=42628742
The new one can.
https://chatgpt.com/share/67e36dee-6694-8010-b337-04f37eeb5c...