Live data from Hacker News

Qwen2.5-VL-32B: Smarter and Lighter

qwenlm.github.io

191–200 of 303 posts

Re: Qwen2.5-VL-32B: Smarter and Lighter

#191
post #152
post #4

Big day for open source Chinese model releases - DeepSeek-v3-0324 came out today too, an updated version of DeepSeek v3 now under an MIT license (previously it was a custom DeepSeek license). https://simonwillison.net/2025/Mar/24/deepseek/

And it still can't answer this: Q: "9.11 and 9.9, which one is larger?" A: "To determine which number is larger between 9.11 and 9.9, let's compare them step by step. Both numbers have the same whole number part: 9. 9.11 has a 1 in the tenths place. 9.9 has a 9 in the tenths place. Since 9 (from 9.9) is greater than 1 (from 9.11), we can conclude that 9.9 is larger than 9.11." "Final Answer: 9.9" I don't think anythi…

What do you think the answer is?

Re: Qwen2.5-VL-32B: Smarter and Lighter

#192
post #134

Earlier quoted context omitted.

That 1% discount feels a bit cheap to me - if it was a 25% or 50% discount I would be much more likely to sign up for it.

We don’t particularly want our customers’ data :)

Yeah, but Openrouter has a 5% surcharge anyway.

Re: Qwen2.5-VL-32B: Smarter and Lighter

#193
post #65

Silly question: how can OpenAI, Claude and all, have a valuation so large considering all the open source models? Not saying they will disappear or be tiny (closed models), but why so so so valuable?

The average user won't self-host a model.

The competition isn't self-hosting. If you can just pick a capable model from any provider inference just turns into a infrastructure/PaaS game -> The majority of the profits will be captured by the cloud providers.

Re: Qwen2.5-VL-32B: Smarter and Lighter

#194
post #152

Earlier quoted context omitted.

And it still can't answer this: Q: "9.11 and 9.9, which one is larger?" A: "To determine which number is larger between 9.11 and 9.9, let's compare them step by step. Both numbers have the same whole number part: 9. 9.11 has a 1 in the tenths place. 9.9 has a 9 in the tenths place. Since 9 (from 9.9) is greater than 1 (from 9.11), we can conclude that 9.9 is larger than 9.11." "Final Answer: 9.9" I don't think anythi…

But that’s correct. 9.9 = 9.90 > 9.11. Seems that it answered the question absolutely correctly.

He's using Semantic versioning/s

Re: Qwen2.5-VL-32B: Smarter and Lighter

#197
post #6

32B is one of my favourite model sizes at this point - large enough to be extremely capable (generally equivalent to GPT-4 March 2023 level performance, which is when LLMs first got really useful) but small enough you can run them on a single GPU or a reasonably well specced Mac laptop (32GB or more).

Are 5090's able to run 32B models?

Re: Qwen2.5-VL-32B: Smarter and Lighter

#198
post #152

Earlier quoted context omitted.

And it still can't answer this: Q: "9.11 and 9.9, which one is larger?" A: "To determine which number is larger between 9.11 and 9.9, let's compare them step by step. Both numbers have the same whole number part: 9. 9.11 has a 1 in the tenths place. 9.9 has a 9 in the tenths place. Since 9 (from 9.9) is greater than 1 (from 9.11), we can conclude that 9.9 is larger than 9.11." "Final Answer: 9.9" I don't think anythi…

I suggest we’ve already now passed what shall be dubbed the jschoe test ;)

I will now refer to this as the jschoe test in my writing and publications as well!

It's interesting to think that maybe one of the most realistic consequences of reaching artificial superintelligence will be when its answers start wildly diverging from human expectations and we think it's being "increasingly wrong".

Re: Qwen2.5-VL-32B: Smarter and Lighter

#199

So today is Qwen. Tomorrow a new SOTA model from Google apparently, R2 next week. We haven't hit the wall yet.

Google's announcements are mostly vaporware anyway. Btw, where is Gemini Ultra 1 ? how about Gemini Ultra 2?

I guess they don’t do ultras anymore, but where was the announcement for it? What other announcement was vaporware?

Re: Qwen2.5-VL-32B: Smarter and Lighter

#200
post #65

Silly question: how can OpenAI, Claude and all, have a valuation so large considering all the open source models? Not saying they will disappear or be tiny (closed models), but why so so so valuable?

Their valuation is not marked to market. We know their previous round valuation, but at this point it is speculative until they go through another round that will mark them again.

That being said, they have a user base and integrations. As long as they stay close or a bit ahead of the Chinese models they'll be fine. If the Chinese models significantly jumps ahead of them, well, then they are pretty much dead. Add open source to the mix and they become history.

Post reply on HN