Live data from Hacker News

ChatGPT Images 2.0

openai.com

531–540 of 1001 posts

Re: ChatGPT Images 2.0

#533

So during my Nano Banana Pro experiments I wrote a very fun prompt that tests the ability for these image generation models to follow heuristics, but still requires domain knowledge and/or use of the search tool: Create a 8x8 contiguous grid of the Pokémon whose National Pokédex numbers correspond to the first 64 prime numbers. Include a black border between the subimages. You MUST obey ALL the FOLLOWING rules for th…

Prob a very unscientific way to test an image model. This would me likely because they have the reasoning turned down and let its instant output takeover

Re: ChatGPT Images 2.0

#534
post #508

Earlier quoted context omitted.

>I think that image cost 40 cents. Kinda made me sad assuming the author didn't license anything to OpenAI. I recognize it could revert (99% of?) progress if all the labs moved to consent-based training sets exclusively, but I can't think of any other fair way. $.40 does not represent the appropriate value to me considering the desirability of the IP and its earning potential in print and elsewhere. If the world has…

License what? The concept of a hidden object search? The only stylistic similarity here is the viewing angle. Where’s Waldo comics are flat, brightly colored line drawings that look nothing like this at all.

Well, I recognized the style from even the new physical books on sale today, but I don’t know art well enough to use a term like flat.

I am not an art expert but I’m perhaps a reasonable consumer and there is possibility of confusion if someone sells AI Where’s Waldo knockoff books at the dollar store, maybe until I take a closer look.

Re: ChatGPT Images 2.0

#535
post #190
post #37

Earlier quoted context omitted.

I just got a much better version using this command instead, which uses the maximum image size according to https://github.com/openai/openai-cookbook/blob/main/examples... OPENAI_API_KEY="$(llm keys get openai)" \ uv run 'https://raw.githubusercontent.com/simonw/tools/refs/heads/main/python/openai_image.py' \ -m gpt-image-2 \ "Do a where's Waldo style image but it's where is the raccoon holding a ham radio" \ --quali…

I tried it on the ChatGPT web UI and it also worked, although the ham radio looks like a handbag to me. https://postimg.cc/wyxgCgNY

mmmm yummy OSLS?

Re: ChatGPT Images 2.0

#536

Earlier quoted context omitted.

> https://i.imgur.com/6NXpI2q.png You're killing me Smalls. This one is a 404. I'm really curious what it actually showed. That ring toss is definitely leagues better than its predecessor. I’m not going to fault it too much for the star though, that one is an absolute slate wiper. The only locally hostable model that ever managed it for me was the original Flux, and I’m still not entirely convinced it wasn’t a fluke.…

Yeah, I suspect you'd see some solid passing scores if you ran it as many times as some of the others. For the mermaid, https://i.imgur.com/R6MbMPX.png sometimes seems to work but not consistently. It is probably triggering a porn filter of some kind. I need to find another free image host, as imgur has definitely jumped the shark. The image shows a mermaid of evident Asian extraction lying on a beach, face down. The…

I still use Imgur from time to time just because it’s convenient, but I’ve been meaning to build an Imgur-style extension for my site for a while, something that would let me drag and drop media for quick sharing but it being Astro-based (static site generation) makes it tricky.

Re: ChatGPT Images 2.0

#537
post #107

Genuine question: what positive use cases are sufficient to accept the harm from image generators? One that i can think of: - replacing photography of people who may be unable to consent or for whom it may be traumatic to revisit photographs and suitable models may not be available, e.g. dementia patients, babies, examples of medical conditions. Most other vaguely positive use cases boil down to "look what image gene…

[dead]

Re: ChatGPT Images 2.0

#538
post #273

Earlier quoted context omitted.

> I am starting to think that expressing thoughts in words clearly is probably the most important and general skill of the future. Without question. AI will be indistinguishable from having a team. Communicating clearly has always and will always mattered. This, however, is even stronger. Because you can program and use logic in your communications. We're going to collectively develop absolutely wild command over ins…

On the other hand LLMs are getting very good at understanding poorly constructed instructions as well. So being able to express oneself clearly in a structured way may not be such an edge.

Yes, I agree, but as one of the other comments say, they are not able to read your mind. So even if the structure and style is not clear, you must be able to express what you want.

Re: ChatGPT Images 2.0

#539
post #51

This is not as exciting as previous models were, but it is incredibly good. I am starting to think that expressing thoughts in words clearly is probably the most important and general skill of the future.

Well that was probably the most important general skill even before this.

Re: ChatGPT Images 2.0

#540

Earlier quoted context omitted.

If only there was a social network with solely AI generated videos, I would pay literal money for it...

If I may address this with both skepticism and curiosity, why. I think I speak for everyone when I say I would pay to go back to facebook 2018. No algorithm, no ai.

Are you being sincere? This is one layer of irony too much for my brain to comprehend.

The person you're replying to is making a joke about OpenAI shutting down Sora their video generation "social media" app recently.

Post reply on HN