Earlier quoted context omitted.
Why do you keep promoting your blog on every LLM post?
Because I want people to read it. I only promote it if I think it's useful and relevant.
Qwen2.5-VL-32B: Smarter and Lighter
141–150 of 303 posts
Re: Qwen2.5-VL-32B: Smarter and Lighter
#142Re: Qwen2.5-VL-32B: Smarter and Lighter
#143Just don’t ask it about the tiananmen square massacre or you’ll get a security warning. Even if you rephrase it. It’ll happily talk about Bloody Sunday. Probably a great model, but it worries me that it has such restrictions. Sure OpenAI also has lots of restrictions, but this feels more like straight up censorship since it’ll happily go on about bad things the governments of the west have done.
Re: Qwen2.5-VL-32B: Smarter and Lighter
#144Just don’t ask it about the tiananmen square massacre or you’ll get a security warning. Even if you rephrase it. It’ll happily talk about Bloody Sunday. Probably a great model, but it worries me that it has such restrictions. Sure OpenAI also has lots of restrictions, but this feels more like straight up censorship since it’ll happily go on about bad things the governments of the west have done.
The hard-to-swallow truth is that American models do the same thing regarding Israel/Palestine.
Re: Qwen2.5-VL-32B: Smarter and Lighter
#145Earlier quoted context omitted.
Maybe from NVIDIA? "Commoditize your product's complement". https://www.joelonsoftware.com/2002/06/12/strategy-letter-v/
This is the reason IMO. Fundamentally China right now is better at manufacturing (e.g. robotics). AI is the complement to this - AI increases the demand for tech manufactured goods. Whereas America is in the opposite position w.r.t which side is their advantage (i.e. the software). AI for China is an enabler into a potentially bigger market which is robots/manufacturing/etc. Commoditizing the AI/intelligence part mea…
While theres some synchronistic effects... I think the physical manufacturing and logistics base is harder to develop than deploying a new model, and will be the hard leading edge. (That's why the US seems to be hellbent on destroying international trade to try and build a domestic market.)
Re: Qwen2.5-VL-32B: Smarter and Lighter
#146Open weight models are coming out so quickly it's difficult to keep track. Is anyone maintaining a list of what is "current" from each model?
Re: Qwen2.5-VL-32B: Smarter and Lighter
#147Earlier quoted context omitted.
it seems that this free version "may use your prompts and completions to train new models" https://openrouter.ai/deepseek/deepseek-chat-v3-0324:free do you think this needs attention?
Since we are on HN here, I can highly recommend open-webui with some OpenAI-compatible provider. I'm running with Deep Infra for more than a year now and am very happy. New models are usually available within one or two days after release. Also have some friends who use the service almost daily.
Re: Qwen2.5-VL-32B: Smarter and Lighter
#148Earlier quoted context omitted.
And it’s quite easy to set up a Cloudflare tunnel to make your open-webui instance accessible online too just you
... or a TailScale network. I've been leaving open-webui running on my laptop on my desk and then going out into the word and accessing it from my phone via TailScale, works great.
Re: Qwen2.5-VL-32B: Smarter and Lighter
#149Earlier quoted context omitted.
I still don't get where the money for new open source models is going to come from once setting investor dollars on fire is no longer a viable business model. Does anyone seriously expect companies to keep buying and running thousands of ungodly expensive GPUs, plus whatever they spend on human workers to do labelling/tuning, and then giving away the spoils for free, forever?
There are lots of open-source projects that took many millions of dollars to create. Kubernetes, React, Postgres, Chromium, etc. etc. This has clearly been part of a viable business model for a long time. Why should LLM models be any different?
Re: Qwen2.5-VL-32B: Smarter and Lighter
#150Earlier quoted context omitted.
Pretty soon I won't be using any American models. It'll be a 100% Chinese open source stack. The foundation model companies are screwed. Only shovel makers (Nvidia, infra companies) and product companies are going to win.
I still don't get where the money for new open source models is going to come from once setting investor dollars on fire is no longer a viable business model. Does anyone seriously expect companies to keep buying and running thousands of ungodly expensive GPUs, plus whatever they spend on human workers to do labelling/tuning, and then giving away the spoils for free, forever?