Live data from Hacker News

Qwen-Image-3.0: Rich Content, Authentic Details, Deep Knowledge

qwen.ai

151–160 of 241 posts

Re: Qwen-Image-3.0: Rich Content, Authentic Details, Deep Knowledge

#152
post #33

To me it feels so weird that people are trying to push these model for online shopping like "here is how this dress/shirt/pants would look on you". But these models will always make the clothes fit your body and show you in flattering light and so on. How the actual garment fits is still as elusive as before these tools

The short-term goal of a tool like this is to sell products. The more ambitious long-term goal is to shift cultural norms, blurring the lines between advertising and reality until the question you're asking is no longer consciously asked. At least, not by average people, and not at the point of purchase.

I find it easy to envision a world, maybe 50 years from now, in which the very concept of "truth in advertising" is viewed as a lost, idyllic fantasy. Something people are nostalgic for, but feel powerless to regain.

Re: Qwen-Image-3.0: Rich Content, Authentic Details, Deep Knowledge

#153

What I can not wrap my head around: How are these models trained? What training mechanism or model architecture provides the glue to go from human text to images? Don't you need to have millions of really descriptively labelled images?

I found this 3b1b guest video on diffusion helpful: https://www.youtube.com/watch?v=iv-5mZ_9CPY

Re: Qwen-Image-3.0: Rich Content, Authentic Details, Deep Knowledge

#155
post #33

To me it feels so weird that people are trying to push these model for online shopping like "here is how this dress/shirt/pants would look on you". But these models will always make the clothes fit your body and show you in flattering light and so on. How the actual garment fits is still as elusive as before these tools

I've been struggling with this question myself. But isn't this a model training/use problem (a.k.a skill issue )? Isn't there a way to make these models be faithful to how people will actually look?

Re: Qwen-Image-3.0: Rich Content, Authentic Details, Deep Knowledge

#156

What I can not wrap my head around: How are these models trained? What training mechanism or model architecture provides the glue to go from human text to images? Don't you need to have millions of really descriptively labelled images?

That's exactly how they do it.

There are ML models that do the reverse and output image to text, which assist quite a lot.

The better the text represents the unique thing in the photo, the better the model understands what that text means.

Re: Qwen-Image-3.0: Rich Content, Authentic Details, Deep Knowledge

#158

They must have trained on GPT Image 1 outputs. The yellow tint is unmistakable. https://qianwen-res.oss-accelerate.aliyuncs.com/Qwen-Image/i... https://qianwen-res.oss-accelerate.aliyuncs.com/Qwen-Image/i... https://qianwen-res.oss-accelerate.aliyuncs.com/Qwen-Image/i... https://qianwen-res.oss-accelerate.aliyuncs.com/Qwen-Image/i...

You clearly haven't met a Chinese RedNote user.

Re: Qwen-Image-3.0: Rich Content, Authentic Details, Deep Knowledge

#159
post #148

Earlier quoted context omitted.

This reddit slop is unbecoming of HN.

[flagged]

> Don’t make me report you to a submod.

This is pathologically outside scope. I don't think I have ever seen somebody threaten somebody else on hn before. The 'don't make me do it to you' abuser trope is next level.

Post reply on HN