It gets the Sally question correct, but it takes more than 100 lines of reasoning. >Sally has three brothers. Each brother has two sisters. How many sisters does sally have? Here is the answer: https://pastebin.com/JP2V92Kh
In fairness it actually works out the correct answer fairly quickly (20 lines, including a false start and correction thereof). It seems to have identified (correctly) that this is a tricky question that it is struggling with so it does a lot of checking.
QwQ: Alibaba's O1-like reasoning LLM
311–320 of 435 posts
Re: QwQ: Alibaba's O1-like reasoning LLM
#312Re: QwQ: Alibaba's O1-like reasoning LLM
#313So western controls on training hardware (hello NVIDIA) seem to have failed. I wonder if there will be any repercussions here.
Re: QwQ: Alibaba's O1-like reasoning LLM
#314Is their repo / model free of any undisclosed telemetry, ie is it purely weights
Is it even possible to embed telemetry into a model itself, as opposed to the runtime environment / program (e.g. Ollama)? I would be disinclined to believe that to be possible, but if anyone knows otherwise, please share.
Re: QwQ: Alibaba's O1-like reasoning LLM
#315So western controls on training hardware (hello NVIDIA) seem to have failed. I wonder if there will be any repercussions here.
Most of the open source models on GitHub, too.
Hailuo, Kling, Vidu, and Hunyuan (posted on Banodoko) blow Sora and Runway out of the water.
China is dominating at this field. And if they begin releasing weights as open source, that'll mean foundation model companies can only bank on the thin facade of product. That's a really good strategy to make sure American AI startups don't achieve escape velocity if they have to fend of dozens of fungible clones.
Re: QwQ: Alibaba's O1-like reasoning LLM
#316So western controls on training hardware (hello NVIDIA) seem to have failed. I wonder if there will be any repercussions here.
Most of the papers in machine learning are coming from China. The vast majority. Most of the open source models on GitHub, too. Hailuo, Kling, Vidu, and Hunyuan (posted on Banodoko) blow Sora and Runway out of the water. China is dominating at this field. And if they begin releasing weights as open source, that'll mean foundation model companies can only bank on the thin facade of product. That's a really good strate…
Re: QwQ: Alibaba's O1-like reasoning LLM
#317So western controls on training hardware (hello NVIDIA) seem to have failed. I wonder if there will be any repercussions here.
Most of the papers in machine learning are coming from China. The vast majority. Most of the open source models on GitHub, too. Hailuo, Kling, Vidu, and Hunyuan (posted on Banodoko) blow Sora and Runway out of the water. China is dominating at this field. And if they begin releasing weights as open source, that'll mean foundation model companies can only bank on the thin facade of product. That's a really good strate…
> that'll mean foundation model companies can only bank on the thin facade of product.
The “facade” of product tested in the real world in the hands of millions or billions is better than thousands of unread/uncited/clique-cited papers using questionable gameable benchmarks.
Re: QwQ: Alibaba's O1-like reasoning LLM
#318Is Alibaba's LLM the "Chinese LLM"? It would appear to have been a U.S.-only game until now. As Eric Schmidt said in the YouTube lecture (that keeps getting pulled down), LLM's have been a rich-companies game.
Re: QwQ: Alibaba's O1-like reasoning LLM
#319Is their repo / model free of any undisclosed telemetry, ie is it purely weights
whatever you load via ollama should be safe, as it only supports uugf and safetensors.
Re: QwQ: Alibaba's O1-like reasoning LLM
#320So western controls on training hardware (hello NVIDIA) seem to have failed. I wonder if there will be any repercussions here.
Or they could be training the models in the states? It’s hard to say since alibaba does R&D in Bellevue as well as Hangzhou.