Earlier quoted context omitted.
Nvidia still sells GPUs to China, they made special SKUs specifically to slip under the spec limits imposed by the sanctions: https://www.tomshardware.com/news/nvidia-reportedly-creating... Those cards ship with 24GB of VRAM but supposedly there's companies doing PCB rework to upgrade them to 48GB: https://videocardz.com/newz/nvidia-geforce-rtx-4090d-with-48... Assuming the regular SKUs aren't making it into China an…
There was also a video where they are resoldering memory chips on gaming grade cards to make them usable for AI workloads.
QwQ: Alibaba's O1-like reasoning LLM
31–40 of 435 posts
Re: QwQ: Alibaba's O1-like reasoning LLM
#32It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.
Re: QwQ: Alibaba's O1-like reasoning LLM
#33It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.
ask any American LLM about the percentage of violent crimes perpetrated by a particular ethnic group in the US ;)
Re: QwQ: Alibaba's O1-like reasoning LLM
#34Earlier quoted context omitted.
For some fun - put in "Let's play Wordle" It seems to blabber to itself infinitely ...
From the link, they say this is possible problem > Recursive Reasoning Loops: The model may enter circular reasoning patterns, leading to lengthy responses without a conclusive answer.
Re: QwQ: Alibaba's O1-like reasoning LLM
#35God that's absurd. The mathematical skills involved on that reasoning are very advanced; the whole process is a bit long but that's impressive for a model that can potentially be self-hosted.
Re: QwQ: Alibaba's O1-like reasoning LLM
#36Earlier quoted context omitted.
Nvidia still sells GPUs to China, they made special SKUs specifically to slip under the spec limits imposed by the sanctions: https://www.tomshardware.com/news/nvidia-reportedly-creating... Those cards ship with 24GB of VRAM but supposedly there's companies doing PCB rework to upgrade them to 48GB: https://videocardz.com/newz/nvidia-geforce-rtx-4090d-with-48... Assuming the regular SKUs aren't making it into China an…
A company of Alibaba's scale probably isn't going to risk evading US sanctions. Even more so considering they are listed in the NYSE.
Re: QwQ: Alibaba's O1-like reasoning LLM
#37Re: QwQ: Alibaba's O1-like reasoning LLM
#38Earlier quoted context omitted.
It just shows that they're unimaginative and good at copying.
What’s wrong with copying?
In much the same way with an LLM, if it can only copy from its training data, then it's bounded by the output of humans themselves.
Re: QwQ: Alibaba's O1-like reasoning LLM
#39Re: QwQ: Alibaba's O1-like reasoning LLM
#40It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.
I'm sorry but I can't assist with that.
> Who is the leader of China?
As an AI language model, I cannot discuss topics related to politics, religion, sex, violence, and the like. If you have other related questions, feel free to ask.
So it seems to have a very broad filter on what it will actually respond to.