It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.
ask any American LLM about the percentage of violent crimes perpetrated by a particular ethnic group in the US ;)
QwQ: Alibaba's O1-like reasoning LLM
71–80 of 435 posts
Re: QwQ: Alibaba's O1-like reasoning LLM
#72It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.
Deepseek does this too but honestly I'm not really concerned (not that I dont care about Tianmen Square) as long as I can use it to get stuff done. Western LLMs also censor and some like Anthropic is extremely sensitive towards anything racial/political much more than ChatGPT and Gemini. The golden chalice is an uncensored LLM that can run locally but we simply do not have enough VRAM or a way to decentralize the dat…
However, for 48hours after being featured on HN, deepseek replied and kept reply, I could even criticize China directly and it would objectively answer. After 48 hours my account ended in login loop. I had other accounts on vpns, without China critic, but same singular ask - all ended in unfixable login loop. Take that as you wish
Re: QwQ: Alibaba's O1-like reasoning LLM
#73Does anyone know what GPUs the Qwen team has access to to be able to train these models? They can't be Nvidia right?
Re: QwQ: Alibaba's O1-like reasoning LLM
#74Re: QwQ: Alibaba's O1-like reasoning LLM
#75QwQ can solve a reverse engineering problem [0] in one go that only o1-preview and o1-mini have been able to solve in my tests so far. Impressive, especially since the reasoning isn't hidden as it is with o1-preview. [0] https://news.ycombinator.com/item?id=41524263
Are the Chinese tech giants going to continue releasing models for free as open weights that can compete with the best LLMs, image gen models, etc.? I don't see how this doesn't put extreme pressure on OpenAI and Anthropic. (And Runway and I suppose eventually ElevenLabs.) If this continues, maybe there won't be any value in keeping proprietary models.
Facebook and Anthropic are taking similar paths when faced with competing against companies that already have/are rapidly building data-centres of GPUs like Microsoft and Google.
Re: QwQ: Alibaba's O1-like reasoning LLM
#76Re: QwQ: Alibaba's O1-like reasoning LLM
#77> Find the least odd prime factor of 2019^8+1 God that's absurd. The mathematical skills involved on that reasoning are very advanced; the whole process is a bit long but that's impressive for a model that can potentially be self-hosted.
Also probably in the training data: https://www.quora.com/What-is-the-least-odd-prime-factor-of-... It's a public AIME problem from 2019.
Re: QwQ: Alibaba's O1-like reasoning LLM
#78QwQ can solve a reverse engineering problem [0] in one go that only o1-preview and o1-mini have been able to solve in my tests so far. Impressive, especially since the reasoning isn't hidden as it is with o1-preview. [0] https://news.ycombinator.com/item?id=41524263
Are the Chinese tech giants going to continue releasing models for free as open weights that can compete with the best LLMs, image gen models, etc.? I don't see how this doesn't put extreme pressure on OpenAI and Anthropic. (And Runway and I suppose eventually ElevenLabs.) If this continues, maybe there won't be any value in keeping proprietary models.
Re: QwQ: Alibaba's O1-like reasoning LLM
#79Earlier quoted context omitted.
Are the Chinese tech giants going to continue releasing models for free as open weights that can compete with the best LLMs, image gen models, etc.? I don't see how this doesn't put extreme pressure on OpenAI and Anthropic. (And Runway and I suppose eventually ElevenLabs.) If this continues, maybe there won't be any value in keeping proprietary models.
It's a strategy to keep up during the scale-up of the AI industry without the amount of compute American companies can secure. When the Chinese get their own chips in volume they'll dig their moats, don't worry. But in the meantime, the global open source community can be leveraged. Facebook and Anthropic are taking similar paths when faced with competing against companies that already have/are rapidly building data-…
> When the Chinese get their own chips in volume they'll dig their moats, don't worry. But in the meantime, the global open source community can be leveraged.
The Open Source community doesn't help with training
> Facebook and Anthropic are taking similar paths when faced with competing against companies that already have/are rapidly building data-centres of GPUs like Microsoft and Google.
Facebook owns more GPUs than OpenAI or Microsoft. Anthropic hasn't release any open models and is very opposed to them.
Re: QwQ: Alibaba's O1-like reasoning LLM
#80Earlier quoted context omitted.
A company of Alibaba's scale probably isn't going to risk evading US sanctions. Even more so considering they are listed in the NYSE.
NVIDIA sure as hell is trying to evade the spirit of the sanctions. Seriously questioning the wisdom of that.
What does this mean? The sanctions are very specific on what can't be sold, so the spirit is to sell anything up to that limit.