Live data from Hacker News

Qwen3-VL

qwen.ai

61–70 of 166 posts

Re: Qwen3-VL

#61

Sadly it still fails the "extra limb" test. I have a few images of animals with an extra limb photoshopped onto them. A dog with an leg coming out of it's stomach, or a cat with two front right legs. Like every other model I have tested, it insists that the animals have their anatomically correct amount of limbs. Even pointing out there is a leg coming from the dogs stomach, it will push back and insist I am confused…

Definitely not a good model for accurately counting limbs on mutant species, then. Might be good at other things that have greater representation in the training set.

Re: Qwen3-VL

#62

Earlier quoted context omitted.

Maybe they just want to see one of the biggest stock bubble pops of all time in the US.

Surprising this is the first time I’ve seen anyone say this out loud.

Because it doesn’t make sense. The reason there’s a bubble is investor belief that AI will unlock tons of value. The reason the bubble is concentrated in silicon and model providers is because investors believe they have the most leverage to monetize this new value in the short term.

If all of that stuff becomes free, the money will just move a few layers up to all of the companies whose cost structure has suddenly been cut dramatically.

There is no commoditization of expensive technology that results in a net loss of market value. It just moves around.

Re: Qwen3-VL

#63

I spent a little time with the thinking model today. It's good. It's not better than GPT5 Pro. It might be better than the smallest GPT 5, though. My current go-to test is to ask the LLM to construct a charging solution for my macbook pro with the model on it, but sadly, I and the pro have been sent to 15th century Florence with no money and no charger. I explain I only have two to three hours of inference time, whic…

I JUST had a very intense dream that there was a catastrophic event that set humanity back massively, to the point that the internet was nonexistent and our laptops suddenly became priceless. The first thought I had was absolutely hating myself for not bothering to download a local LLM. A local LLM at the level of qwen is enough to massively jump start civilization.

Re: Qwen3-VL

#64

Sadly it still fails the "extra limb" test. I have a few images of animals with an extra limb photoshopped onto them. A dog with an leg coming out of it's stomach, or a cat with two front right legs. Like every other model I have tested, it insists that the animals have their anatomically correct amount of limbs. Even pointing out there is a leg coming from the dogs stomach, it will push back and insist I am confused…

I wonder if you used their image editing feature if it would insist on “correcting” the number of limbs even if you asked for unrelated changes.

Re: Qwen3-VL

#65
post #60

China is winning the hearts of developers in this race so far. At least, they won mine already.

Arguably they’ve already won. Check the names at the top the next time you see a paper from an American company, a lot of them are Chinese.

you can’t tell if someone is American or Chinese by looking at their name

I actually claim something even stronger, which is it’s what’s in your heart that really determines if you’re American :)

Re: Qwen3-VL

#66
post #35

Earlier quoted context omitted.

> Take the core technology and just optimize, optimize, optimize for 10x the cost/efficiency. As simple as that. Super impressive. This "just" is incorrect. The Qwen team invented things like DeepStack https://arxiv.org/abs/2406.04334 (Also I hate this "The Chinese" thing. Do we say "The British" if it came from a DeepMind team in the UK? Or what if there are Chinese born US citizens working in Paris for Mistral? Giv…

> Also I hate this "The Chinese" thing to me it was positive assessment, I adore their craftsmanship and persistence in moving forward for long period of time.

It erases the individuals doing the actual research by viewing Chinese people as a monolith.

Re: Qwen3-VL

#67

Roughly 1/10 the cost of Opus 4.1, 1/2 the cost of Sonnet 4 on per token inference basis. Impressive. I'd love to see a fast (groq style) version of this served. I wonder if the architecture is amenable.

Isnt it a 3x rate difference? 0.7$ for Qwen3-VL vs 3$ for Sonnet 4?

Re: Qwen3-VL

#68
Imagine the demand for a 128GB/256GB/512GB unified memory stuffed hardware linux box shipping with Qwen models already up and running.

Although I´m agAInst steps towards AGI, it feels safer to have these things running locally and disconnected from each other, than some giant GW cloud agentic data centers connected to everyone and everything.

Re: Qwen3-VL

#69

Imagine the demand for a 128GB/256GB/512GB unified memory stuffed hardware linux box shipping with Qwen models already up and running. Although I´m agAInst steps towards AGI, it feels safer to have these things running locally and disconnected from each other, than some giant GW cloud agentic data centers connected to everyone and everything.

I bought an GMKtec evo 2 that is a 128 GB unified memory system. Strong recommend.

Re: Qwen3-VL

#70

I spent a little time with the thinking model today. It's good. It's not better than GPT5 Pro. It might be better than the smallest GPT 5, though. My current go-to test is to ask the LLM to construct a charging solution for my macbook pro with the model on it, but sadly, I and the pro have been sent to 15th century Florence with no money and no charger. I explain I only have two to three hours of inference time, whic…

Funny enough, I did a little bit of ChatGPT-assisted research into a loosely similar scenario not too long ago. LPT: if you happen to know in advance that you'll be in Renaissance Florence, make sure to pack as many synthetic diamonds as you can afford.
Post reply on HN