Live data from Hacker News

4o Image Generation

openai.com

491–500 of 629 posts

Re: 4o Image Generation

#491
I enjoy trying to break these models. I come up with prompts that are uncommon but valid. I want to see how well they handle data not in their training set. For image generation I like to use “ Generate an image of a woman on vacation in the Caribbean, lying down on the beach without sunglasses, her eyes open.”

Re: 4o Image Generation

#492
post #104

What's important about this new type of image generation that's happening with tokens rather than with diffusion, is that this is effectively reasoning in pixel space. Example: Ask it to draw a notepad with an empty tic-tac-toe, then tell it to make the first move, then you make a move, and so on. You can also do very impressive information-conserving translations, such as changing the drawing style, but also stuff l…

> truly generative UI, where the model produces the next frame of the app I built this exact thing last month, demo: https://universal.oroborus.org (not viable on phone for this demo, fine on tablet or computer) Also see discussion and code at: http://github.com/snickell/universal I wasn't really planning to share/release it today, but, heck, why not. I started with bitmap-style generative image models, but because t…

Wonderful, good job! Reminds me of https://arstechnica.com/information-technology/2022/12/opena...

Re: 4o Image Generation

#494
post #104

What's important about this new type of image generation that's happening with tokens rather than with diffusion, is that this is effectively reasoning in pixel space. Example: Ask it to draw a notepad with an empty tic-tac-toe, then tell it to make the first move, then you make a move, and so on. You can also do very impressive information-conserving translations, such as changing the drawing style, but also stuff l…

> truly generative UI, where the model produces the next frame of the app I built this exact thing last month, demo: https://universal.oroborus.org (not viable on phone for this demo, fine on tablet or computer) Also see discussion and code at: http://github.com/snickell/universal I wasn't really planning to share/release it today, but, heck, why not. I started with bitmap-style generative image models, but because t…

I’m a bit late here - but I’m the COO of OpenRouter and would love to help out with some additional credits and share the project. It’s very cool and more people could be able to check it out. Send me a note. My email is cc at OpenRouter.ai

Re: 4o Image Generation

#495
post #408

Earlier quoted context omitted.

I am using Plus from Australia, and while I am not getting a full glass, nor am I getting a half full glass. The glass I'm getting is half empty.

That's funny. HN hates funny. Enjoy your shadowban.

Yeah. I understand that this site doesn’t want to become Reddit, but it really has an allergy to comedy, it’s sad. God forbid you use sarcasm, half the people here won’t understand it and the other half will say it’s not appropriate for healthy discussion…

Re: 4o Image Generation

#496

OpenAI's livestream of GPT-4o Image Generation shows that it is slowwwwwwwwww (maybe 30 seconds per image, which Sam Altman had to spin "it's slow but the generated images are worth it"). Instead of using a diffusion approach, it appears to be generating the image tokens and decoding them akin to the original DALL-E ( https://openai.com/index/dall-e/ ), which allows for streaming partial generations from top to botto…

If you look at the examples given, this is the first time I've felt like AI generated images have passed the uncanny valley. The results are ground breaking in my opinion. How much longer until an AI can generate 30 successive images together and make an ultra realistic movie?

One day you’ll just give it a script and get a movie out

Re: 4o Image Generation

#497

Ran through some of my relatively complex prompts combined with using pure text prompts as the de-facto means of making adjustments to the images (in contrast to using something like img2img / inpainting / etc.) https://mordenstar.com/blog/chatgpt-4o-images It's definitely impressive though once again fell flat on the ability to render a 9-pointed star.

Fantastic prompts!

Re: 4o Image Generation

#498
post #205

Anyone else frightened by this? Seeing meant believing, and now that isnt the case anymore...

it's fine, there will be new jobs.

Yes, but not in time to save the people who were in the old jobs. Plus retraining.

over 10 years it might even out, if your lucky (historically its taken much longer) but 10 years is a long time to wait in your career.

Re: 4o Image Generation

#500
post #393

My experience with these announcements is that they're cherry picking the best results from a maybe several hundred or a thousand prompts. I'm not saying that it's not true, it's just "wait and see" before you take their word as gold. I think MS's claim on their quantum computing breakthrough is the latest form of this.

> My experience with these announcements is that they're cherry picking the best results from a maybe several hundred or a thousand prompt

just tried it, prompt adherence and quality is... exactly what they said, it extremely impressive

Post reply on HN