QwQ-32B: Embracing the Power of Reinforcement Learning
qwenlm.github.io
QwQ-32B: Embracing the Power of Reinforcement Learning
1–10 of 178 posts
Re: QwQ-32B: Embracing the Power of Reinforcement Learning
#2This is insane matching deepseek but 20x smaller?
Re: QwQ-32B: Embracing the Power of Reinforcement Learning
#3[deleted]
Re: QwQ-32B: Embracing the Power of Reinforcement Learning
#4No bad.
I have tried it in a current project (Online Course) where Deepseek and Gemini have done a good job with a "stable" prompt and my impression is: -Somewhat simplified but original answers
We will have to keep an eye on it
Re: QwQ-32B: Embracing the Power of Reinforcement Learning
#5Chinese strategy is open-source software part and earn on robotics part. And, They are already ahead of everyone in that game.
These things are pretty interesting as they are developing. What US will do to retain its power?
BTW I am Indian and we are not even in the race as country. :(
Re: QwQ-32B: Embracing the Power of Reinforcement Learning
#6This is ridiculous. 32B and beating deepseek and o1. And yet I'm trying it out and, yeah, it seems pretty intelligent...
Remember when models this size could just about maintain a conversation?
Re: QwQ-32B: Embracing the Power of Reinforcement Learning
#7To test: https://chat.qwen.ai/ and select Qwen2.5-plus, then toggle QWQ.
Re: QwQ-32B: Embracing the Power of Reinforcement Learning
#8Nice. Hard to tell whether it's really on a par with o1 or R1, but it's definitely very impressive for a 32B model.
Re: QwQ-32B: Embracing the Power of Reinforcement Learning
#9actually insane how small the model is. they are only going to get better AND smaller. wild times
Re: QwQ-32B: Embracing the Power of Reinforcement Learning
#10This is insane matching deepseek but 20x smaller?
Roughly the same number of active parameters as R1 is a mixture-of-experts model. Still extremely impressive, but not unbelievable.