4o Image Generation
491–500 of 629 posts
Re: 4o Image Generation
#492What's important about this new type of image generation that's happening with tokens rather than with diffusion, is that this is effectively reasoning in pixel space. Example: Ask it to draw a notepad with an empty tic-tac-toe, then tell it to make the first move, then you make a move, and so on. You can also do very impressive information-conserving translations, such as changing the drawing style, but also stuff l…
> truly generative UI, where the model produces the next frame of the app I built this exact thing last month, demo: https://universal.oroborus.org (not viable on phone for this demo, fine on tablet or computer) Also see discussion and code at: http://github.com/snickell/universal I wasn't really planning to share/release it today, but, heck, why not. I started with bitmap-style generative image models, but because t…
Re: 4o Image Generation
#493I couldn't find anything on the pricing page.
Re: 4o Image Generation
#494What's important about this new type of image generation that's happening with tokens rather than with diffusion, is that this is effectively reasoning in pixel space. Example: Ask it to draw a notepad with an empty tic-tac-toe, then tell it to make the first move, then you make a move, and so on. You can also do very impressive information-conserving translations, such as changing the drawing style, but also stuff l…
> truly generative UI, where the model produces the next frame of the app I built this exact thing last month, demo: https://universal.oroborus.org (not viable on phone for this demo, fine on tablet or computer) Also see discussion and code at: http://github.com/snickell/universal I wasn't really planning to share/release it today, but, heck, why not. I started with bitmap-style generative image models, but because t…
Re: 4o Image Generation
#495Earlier quoted context omitted.
I am using Plus from Australia, and while I am not getting a full glass, nor am I getting a half full glass. The glass I'm getting is half empty.
That's funny. HN hates funny. Enjoy your shadowban.
Re: 4o Image Generation
#496OpenAI's livestream of GPT-4o Image Generation shows that it is slowwwwwwwwww (maybe 30 seconds per image, which Sam Altman had to spin "it's slow but the generated images are worth it"). Instead of using a diffusion approach, it appears to be generating the image tokens and decoding them akin to the original DALL-E ( https://openai.com/index/dall-e/ ), which allows for streaming partial generations from top to botto…
If you look at the examples given, this is the first time I've felt like AI generated images have passed the uncanny valley. The results are ground breaking in my opinion. How much longer until an AI can generate 30 successive images together and make an ultra realistic movie?
Re: 4o Image Generation
#497Ran through some of my relatively complex prompts combined with using pure text prompts as the de-facto means of making adjustments to the images (in contrast to using something like img2img / inpainting / etc.) https://mordenstar.com/blog/chatgpt-4o-images It's definitely impressive though once again fell flat on the ability to render a 9-pointed star.
Re: 4o Image Generation
#498Anyone else frightened by this? Seeing meant believing, and now that isnt the case anymore...
it's fine, there will be new jobs.
over 10 years it might even out, if your lucky (historically its taken much longer) but 10 years is a long time to wait in your career.
Re: 4o Image Generation
#499Re: 4o Image Generation
#500My experience with these announcements is that they're cherry picking the best results from a maybe several hundred or a thousand prompts. I'm not saying that it's not true, it's just "wait and see" before you take their word as gold. I think MS's claim on their quantum computing breakthrough is the latest form of this.
just tried it, prompt adherence and quality is... exactly what they said, it extremely impressive