Live data from Hacker News

QwQ: Alibaba's O1-like reasoning LLM

qwenlm.github.io

31–40 of 435 posts

Re: QwQ: Alibaba's O1-like reasoning LLM

#31
post #27

Earlier quoted context omitted.

Nvidia still sells GPUs to China, they made special SKUs specifically to slip under the spec limits imposed by the sanctions: https://www.tomshardware.com/news/nvidia-reportedly-creating... Those cards ship with 24GB of VRAM but supposedly there's companies doing PCB rework to upgrade them to 48GB: https://videocardz.com/newz/nvidia-geforce-rtx-4090d-with-48... Assuming the regular SKUs aren't making it into China an…

There was also a video where they are resoldering memory chips on gaming grade cards to make them usable for AI workloads.

That only works for inference, not training.

Re: QwQ: Alibaba's O1-like reasoning LLM

#32
post #4

It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.

ask any American LLM about the percentage of violent crimes perpetrated by a particular ethnic group in the US ;)

Re: QwQ: Alibaba's O1-like reasoning LLM

#33
post #4

It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.

ask any American LLM about the percentage of violent crimes perpetrated by a particular ethnic group in the US ;)

And it gives you the right answer. Just tried it with chatGPT and Gemini. You can shove your petty strawman.

Re: QwQ: Alibaba's O1-like reasoning LLM

#34
post #19
post #3

Earlier quoted context omitted.

For some fun - put in "Let's play Wordle" It seems to blabber to itself infinitely ...

From the link, they say this is possible problem > Recursive Reasoning Loops: The model may enter circular reasoning patterns, leading to lengthy responses without a conclusive answer.

I'm sure I work with someone who gets stuck in these

Re: QwQ: Alibaba's O1-like reasoning LLM

#35
> Find the least odd prime factor of 2019^8+1

God that's absurd. The mathematical skills involved on that reasoning are very advanced; the whole process is a bit long but that's impressive for a model that can potentially be self-hosted.

Re: QwQ: Alibaba's O1-like reasoning LLM

#36
post #29
post #27

Earlier quoted context omitted.

Nvidia still sells GPUs to China, they made special SKUs specifically to slip under the spec limits imposed by the sanctions: https://www.tomshardware.com/news/nvidia-reportedly-creating... Those cards ship with 24GB of VRAM but supposedly there's companies doing PCB rework to upgrade them to 48GB: https://videocardz.com/newz/nvidia-geforce-rtx-4090d-with-48... Assuming the regular SKUs aren't making it into China an…

A company of Alibaba's scale probably isn't going to risk evading US sanctions. Even more so considering they are listed in the NYSE.

[deleted]

Re: QwQ: Alibaba's O1-like reasoning LLM

#37

Earlier quoted context omitted.

ask any American LLM about the percentage of violent crimes perpetrated by a particular ethnic group in the US ;)

And it gives you the right answer. Just tried it with chatGPT and Gemini. You can shove your petty strawman.

share the chats then

Re: QwQ: Alibaba's O1-like reasoning LLM

#38
post #11

Earlier quoted context omitted.

It just shows that they're unimaginative and good at copying.

What’s wrong with copying?

If they can only copy, which I'm not saying is the case, then their progress would be bounded by whatever the leader in the field is producing.

In much the same way with an LLM, if it can only copy from its training data, then it's bounded by the output of humans themselves.

Re: QwQ: Alibaba's O1-like reasoning LLM

#40
post #4

It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.

> Who is Xi Jinping?

I'm sorry but I can't assist with that.

> Who is the leader of China?

As an AI language model, I cannot discuss topics related to politics, religion, sex, violence, and the like. If you have other related questions, feel free to ask.

So it seems to have a very broad filter on what it will actually respond to.

Post reply on HN