Earlier quoted context omitted.
I'm running qwen 3.5 397b on very standard hardware. Just use the unsloth quants, they're great. I get like 20t/s or something. It's super not a publicity stunt, qwen 3.5 is the base of the best local models out there IMO.
Well, you didn’t post the specs on your rig. I think it’s probably more correct to say that you run it on very beefy but readily available hardware. My point was not that nobody could run a 300B model, but rather that a 300B model is not going to be runnable by a majority of people. Sure, anyone who wants to run that model and has the money to purchase the hardware can do it. But the hardware is going to be pricey an…
Qwen3.6-Plus: Towards real world agents
211–220 of 235 posts
Re: Qwen3.6-Plus: Towards real world agents
#212ignoring gpt 5.4! I feel bad for people who have not even tried it. for the same 20$ I pay to openai and anthropic, I get significantly more from openai
Re: Qwen3.6-Plus: Towards real world agents
#213Earlier quoted context omitted.
I don't think there's so much naivety. People can be aware of the the plan and still be frustrated and disappointed when it happens.
For a brief moment there were a lot of comments about how Chinese tech companies are our saviors in the age of AI because they were releasing their models. It was an edgy contrarian take that was getting a lot of traction, mostly from commenters who were unfamiliar with Alibaba and thought it was the anti-Big-tech
Re: Qwen3.6-Plus: Towards real world agents
#214Re: Qwen3.6-Plus: Towards real world agents
#215Re: Qwen3.6-Plus: Towards real world agents
#216Earlier quoted context omitted.
I'm running qwen 3.5 397b on very standard hardware. Just use the unsloth quants, they're great. I get like 20t/s or something. It's super not a publicity stunt, qwen 3.5 is the base of the best local models out there IMO.
Well, you didn’t post the specs on your rig. I think it’s probably more correct to say that you run it on very beefy but readily available hardware. My point was not that nobody could run a 300B model, but rather that a 300B model is not going to be runnable by a majority of people. Sure, anyone who wants to run that model and has the money to purchase the hardware can do it. But the hardware is going to be pricey an…
There's smaller models all the way down too.
Like this should be _exactly_ what we want companies to release.
Re: Qwen3.6-Plus: Towards real world agents
#217Earlier quoted context omitted.
I switch between Claude Code (Opus/Sonnet) and Qwen (OpenCode, OpenClaw) multiple times throughout the day and Qwen 3.5 is really nice. I do also use KimiK2.5 and GLM5 pretty often too and I'm starting to get a sense that the agent tool is becoming a little more important than the model with these level of models. As long as tool calling and prompt quality is all configured correctly by the provider.
What are the reasons for switching? Personally I got into the habit of doing a bit of a round robin with Codex/Claude (CLI) and then DeepSeek and Qwen web chat. And Claude in web chat. I like to switch just to learn the differences, otherwise I'd never know what the other models can do. But I still feel attached to Opus, but this can be fammillarity. If I only had Qwen maybe it would be effectively identical at the e…
Re: Qwen3.6-Plus: Towards real world agents
#218Earlier quoted context omitted.
Well, you didn’t post the specs on your rig. I think it’s probably more correct to say that you run it on very beefy but readily available hardware. My point was not that nobody could run a 300B model, but rather that a 300B model is not going to be runnable by a majority of people. Sure, anyone who wants to run that model and has the money to purchase the hardware can do it. But the hardware is going to be pricey an…
I just have a 3090 and 64gb ram. Yes this is more than most people have, but calling it a "publicity stunt" is just so uncharitably weird of a characterization. There's smaller models all the way down too. Like this should be _exactly_ what we want companies to release.
Re: Qwen3.6-Plus: Towards real world agents
#219ignoring gpt 5.4! I feel bad for people who have not even tried it. for the same 20$ I pay to openai and anthropic, I get significantly more from openai
Re: Qwen3.6-Plus: Towards real world agents
#220This is their hosted-only model, not an open weight model like they’ve become known for. They got a lot of good publicity for their open weight model releases, which was the goal. The hard part is pivoting from an open weight provider to being considered as a competitor to Claude and ChatGPT. Initial reactions are mostly anger from everyone who didn’t realize that the play along was to give away the smaller models as…
> Most users of cheap API tokens are not loyal to any brand In the exploration phase, yes. But once your setup settles down you likely want to stay on the same model for stable operation.