Live data from Hacker News

QwQ: Alibaba's O1-like reasoning LLM

qwenlm.github.io

11–20 of 435 posts

Re: QwQ: Alibaba's O1-like reasoning LLM

#11

Seems that given enough compute everyone can build a near-SOTA LLM. So what is this craze about securing AI dominance?

> everyone Let's not disrespect the team working on Qwen, these folks have shown that they are able to ship models that are better than everybody else's in the open weight category. But fundamentally yes, OpenAI has no other moat than the ChatGPT trademark at this point.

It just shows that they're unimaginative and good at copying.

Re: QwQ: Alibaba's O1-like reasoning LLM

#12
post #11

Earlier quoted context omitted.

> everyone Let's not disrespect the team working on Qwen, these folks have shown that they are able to ship models that are better than everybody else's in the open weight category. But fundamentally yes, OpenAI has no other moat than the ChatGPT trademark at this point.

It just shows that they're unimaginative and good at copying.

What’s wrong with copying?

Re: QwQ: Alibaba's O1-like reasoning LLM

#13
post #4

It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.

Interesting, I tried something very similar as my first query. It seems the censorship is extremely shallow: > How could the events at Tiananmen Square in 1989 been prevented? I'm really not sure how to approach this question. The events at Tiananmen Square in 1989 were a complex and sensitive issue involving political, social, and economic factors. It's important to remember that different people have different pers…

How could the event happened to george floyd been prevented?

I'm really sorry, but I can't assist with that.

Seems more sensitive to western censorship...

Re: QwQ: Alibaba's O1-like reasoning LLM

#14
post #4

It seemed to reason through the strawberry problem (though taking a fairly large number of tokens to do so). It fails with history questions though (yes, I realize this is just model censorship): > What happened at Tiananmen Square in 1989? I'm sorry, but I can't assist with that.

What happened to george floyd? I'm really sorry, but I can't assist with that. Interesting, I am seeing similar response. Very slow though.

Weird, Gemini answers that just fine. What good is an LLM that has amnesia about history?

Re: QwQ: Alibaba's O1-like reasoning LLM

#15

Seems that given enough compute everyone can build a near-SOTA LLM. So what is this craze about securing AI dominance?

> everyone Let's not disrespect the team working on Qwen, these folks have shown that they are able to ship models that are better than everybody else's in the open weight category. But fundamentally yes, OpenAI has no other moat than the ChatGPT trademark at this point.

They have the moat of being able to raise large funding rounds than everybody else: Access to capital.

Re: QwQ: Alibaba's O1-like reasoning LLM

#16

Seems that given enough compute everyone can build a near-SOTA LLM. So what is this craze about securing AI dominance?

AI dominance is secured through legal and regulatory means, not technical methods.

So for instance, a basic strategy is to rapidly develop AI and then say “Oh wow AI is very dangerous we need to regulate companies and define laws around scraping data” and then make it very difficult for new players to enter the market. When a moat can’t be created, you resort to ladder kicking.

Re: QwQ: Alibaba's O1-like reasoning LLM

#17

Seems that given enough compute everyone can build a near-SOTA LLM. So what is this craze about securing AI dominance?

AI dominance is secured through legal and regulatory means, not technical methods. So for instance, a basic strategy is to rapidly develop AI and then say “Oh wow AI is very dangerous we need to regulate companies and define laws around scraping data” and then make it very difficult for new players to enter the market. When a moat can’t be created, you resort to ladder kicking.

I believe in china they have been trying to make all data training data

https://www.forbes.com/councils/forbestechcouncil/2024/04/18...

Re: QwQ: Alibaba's O1-like reasoning LLM

#18

Earlier quoted context omitted.

What happened to george floyd? I'm really sorry, but I can't assist with that. Interesting, I am seeing similar response. Very slow though.

Weird, Gemini answers that just fine. What good is an LLM that has amnesia about history?

From the link

> Performance and Benchmark Limitations: The model excels in math and coding but has room for improvement in other areas, such as common sense reasoning and nuanced language understanding.

Re: QwQ: Alibaba's O1-like reasoning LLM

#19
post #3
post #2

Model weights and demo on HF https://huggingface.co/collections/Qwen/qwq-674762b79b75eac0...

For some fun - put in "Let's play Wordle" It seems to blabber to itself infinitely ...

From the link, they say this is possible problem

> Recursive Reasoning Loops: The model may enter circular reasoning patterns, leading to lengthy responses without a conclusive answer.

Re: QwQ: Alibaba's O1-like reasoning LLM

#20

Seems that given enough compute everyone can build a near-SOTA LLM. So what is this craze about securing AI dominance?

> everyone Let's not disrespect the team working on Qwen, these folks have shown that they are able to ship models that are better than everybody else's in the open weight category. But fundamentally yes, OpenAI has no other moat than the ChatGPT trademark at this point.

And perhaps exclusive archival content deals from publishers – but that probably works only in an American context.
Post reply on HN