Is Alibaba's LLM the "Chinese LLM"? It would appear to have been a U.S.-only game until now. As Eric Schmidt said in the YouTube lecture (that keeps getting pulled down), LLM's have been a rich-companies game.
you only think that because you haven’t been paying close attention qwen, deepseek, yi - there have been a number of high quality, open chinese competitors
QwQ: Alibaba's O1-like reasoning LLM
301–310 of 435 posts
Re: QwQ: Alibaba's O1-like reasoning LLM
#302I can't wait for Ebay to release theirs
Re: QwQ: Alibaba's O1-like reasoning LLM
#303Re: QwQ: Alibaba's O1-like reasoning LLM
#304Earlier quoted context omitted.
I haven’t ran QWQ yet, but it’s a 32B. So about 20GB RAM with Q4 quant. Closer to 25GB for the 4_K_M one. You can wait for a day or so for the quantized GGUFs to show up (we should see the Q4 in the next hour or so). I personally use Ollama on an MacBook Pro. It usually takes a day or two for it to show up. Any M series MacBook with 32GB+ of RAM will run this.
On Macbooks with Apple Silicon consider MLX models from MLX community: https://huggingface.co/collections/mlx-community/qwq-32b-pre... For a GUI, LM Studio 0.3.x is iterating MLX support: https://lmstudio.ai/beta-releases When searching in LM Studio, you can narrow search to the mlx-community.
also I didn't install a beta and mine says i'm using 3.5 which is what the beta also says. is there a difference right now between the beta and the release version?
Re: QwQ: Alibaba's O1-like reasoning LLM
#305Re: QwQ: Alibaba's O1-like reasoning LLM
#306Earlier quoted context omitted.
Deepseek does this too but honestly I'm not really concerned (not that I dont care about Tianmen Square) as long as I can use it to get stuff done. Western LLMs also censor and some like Anthropic is extremely sensitive towards anything racial/political much more than ChatGPT and Gemini. The golden chalice is an uncensored LLM that can run locally but we simply do not have enough VRAM or a way to decentralize the dat…
Ask Anthropic whether the USA has ever comitted war crimes, and it said "yes" and listed ten, including the My Lai Massacre in Vietname and Abu Graib. The political censorship is not remotely comparable.
Re: QwQ: Alibaba's O1-like reasoning LLM
#307Is o1 even that good? It's doesn't even rank first on LMArena..
My only real problem with o1 is that it's ridiculously expensive, to the point that it makes no sense to use it for actual code. In architect mode, however, you can keep the costs under control as there are far fewer input/output tokens.
Re: QwQ: Alibaba's O1-like reasoning LLM
#308Re: QwQ: Alibaba's O1-like reasoning LLM
#309I am right now playing with it running it locally using ollama. It is a 19GB download and it runs nicely on a nvidia A100 GPU. https://ollama.com/library/qwq
Re: QwQ: Alibaba's O1-like reasoning LLM
#310Earlier quoted context omitted.
you only think that because you haven’t been paying close attention qwen, deepseek, yi - there have been a number of high quality, open chinese competitors
And AI21, which is Israeli