Live data from Hacker News

ChatGPT Images 2.0

openai.com

551–560 of 1001 posts

Re: ChatGPT Images 2.0

#551

Earlier quoted context omitted.

The faces...that's nice that it turned a kid's book into an abomination

By image generation standards this is a ridiculously good result. No surprise that people instantly find the new limits, but they are new limits.

But it's also straight up plagiarism and still ridiculously bad on so many levels.

Re: ChatGPT Images 2.0

#552
post #438

Are camera manufacturers working on signed images? That seems like the only way our trust in any digital media doesn't collapse entirely.

Ultimately even with that tech, you can still take a photo of an AI generated scene. Maybe coupled with geolocation data in the signature or something it might work.

Any thoughts on attempted multiple camera/360 camera solutions? Can make it cost prohibitive to generate exceptional fakes… for a little while

Kind of like showing the proctor around your room with your webcam before starting the exam.

I think legacy media stands a chance at coming back as long as they maintain a reputation of deeply verifying images, not being fooled.

Re: ChatGPT Images 2.0

#553

Are camera manufacturers working on signed images? That seems like the only way our trust in any digital media doesn't collapse entirely.

Signed images don’t get you much. You can just hardwire the image sensor to a computer and sign raw pixels.

Is the situation brighter for a company who owns the hardware and the software, for Apple?

Taking a picture of an AI generated image aside, theoretically could Apple attest to origin of photos taken in the native camera app and uploaded to iCloud?

Fascinating, by the way, thank you!

Re: ChatGPT Images 2.0

#554
A great technical achievement, for sure, but this is kind of the moment where it enters uncanny valley to me. The promo reel on the website makes it feel like humans doing incredible things (background music intentionally evokes that emotion), but it's a slideshow of computer generatated images attempting to replicate the amazing things that humans do. It's just crazy to look at those images and have to consciously remind myself - nobody made this, this photographed place and people do not exist, no human participated in this photo, no human traced the lines of this comic, no human designer laid out the text in this image. This is a really clever amalgamation machine of human-based inputs. Uncanny valley.

Re: ChatGPT Images 2.0

#555

And here I was proud of myself, having taught my mom and her friends how to discern real from fakes they get on WhatsApp groups. Another even more powerful tool for scammers. I'm taking a break.

I told my mom not to believe anything unless she trusts the source. The way people always did with text.

Re: ChatGPT Images 2.0

#556

Earlier quoted context omitted.

I think we are just going to have to accept that realistic images can be easily fabricated now. Seeing is not believing anymore, and I don't think SynthID or anything like it can restore that trust in images.

It's going to mess up accountability. Some politician will be recorded doing something & he'll have his people release a thousand photos/videos of him doing crimes. And they'll say, look, it's a smear campaign. This is just one stupid example, but people will have better schemes. Also global coordinated releases of fake content and hypertargeted possibly abusive content. Virtual kidnappings will take off, automated &…

Some politician will be recorded doing something & he'll have his people release a thousand photos/videos of him doing crimes. And they'll say, look, it's a smear campaign.

And his enemies will do the same, hopefully resulting in less blind trust for everyone in the population, which can only be a good thing.

Re: ChatGPT Images 2.0

#557

So during my Nano Banana Pro experiments I wrote a very fun prompt that tests the ability for these image generation models to follow heuristics, but still requires domain knowledge and/or use of the search tool: Create a 8x8 contiguous grid of the Pokémon whose National Pokédex numbers correspond to the first 64 prime numbers. Include a black border between the subimages. You MUST obey ALL the FOLLOWING rules for th…

banana Pro gets the logic and punts on the art; gpt-2-image gets the art and punts on the logic. Feels like instruction-following and creativity sit on opposite ends of the same slider.

Re: ChatGPT Images 2.0

#558
post #464

Earlier quoted context omitted.

[flagged]

What would make the prompt a better actual evaluation in your judgement?

Not focusing on pokemon for a start. Maybe use something more people can recognize and evaluate. I have zero knowledge of pokemon, I see it as a niche thing for ultra-nerdy people, and not something everyone is familiar with. Nothing about that test can be evaluated by anyone but a pokemon expert. Sorry, but pokemon isn't as mainstream as some people might think it is.

Re: ChatGPT Images 2.0

#559
post #209
post #107

Genuine question: what positive use cases are sufficient to accept the harm from image generators? One that i can think of: - replacing photography of people who may be unable to consent or for whom it may be traumatic to revisit photographs and suitable models may not be available, e.g. dementia patients, babies, examples of medical conditions. Most other vaguely positive use cases boil down to "look what image gene…

There are many use-cases outside of spam and slop. For example, take a picture of your garden. Ask chatgpt to give you ideas how to improve it and a step by visual guide. Anything that can be expressed visually is effectively target for this technology - this covers pretty much everything.

That's a multimodal model with text output, I think GP is asking about image generators.

Re: ChatGPT Images 2.0

#560
post #190
post #37

Earlier quoted context omitted.

I just got a much better version using this command instead, which uses the maximum image size according to https://github.com/openai/openai-cookbook/blob/main/examples... OPENAI_API_KEY="$(llm keys get openai)" \ uv run 'https://raw.githubusercontent.com/simonw/tools/refs/heads/main/python/openai_image.py' \ -m gpt-image-2 \ "Do a where's Waldo style image but it's where is the raccoon holding a ham radio" \ --quali…

I tried it on the ChatGPT web UI and it also worked, although the ham radio looks like a handbag to me. https://postimg.cc/wyxgCgNY

Nice, enjoyed the image as someone who has been to the events. But also easy raccoon placement :)
Post reply on HN