Live data from Hacker News

Qwen3.6-Plus: Towards real world agents

qwen.ai

141–150 of 235 posts

Re: Qwen3.6-Plus: Towards real world agents

#141

This is their hosted-only model, not an open weight model like they’ve become known for. They got a lot of good publicity for their open weight model releases, which was the goal. The hard part is pivoting from an open weight provider to being considered as a competitor to Claude and ChatGPT. Initial reactions are mostly anger from everyone who didn’t realize that the play along was to give away the smaller models as…

I use different models in production and model's "personality" as in tendency to not go off script, not consume gazillions of tokens recursively, follow instructions etc, are more relevant than "brute" power which is okayish as a metric for agentic coding on generous token plans.

Chinese models are very competitive in that regard, you'll often look at 70-90% price reduction at the same quality.

Re: Qwen3.6-Plus: Towards real world agents

#142
post #63

Earlier quoted context omitted.

I think it’s more the principle of deception that upsets people. Imagine if Apple released a new iPhone and publicly compared its specs to some previous gen Android. It’s not in good faith.

They compared their M-series chips to older Intel Macs for a while, likely to target users who were still on Intel chips. If they released a lower cost iPhone and compared it to a previous gen Android I could see the reasoning for it. It's not deception if it's a valid comparison and people just fail to understand what's being compared. Now, is it mildly deceptive because all of the companies using incredibly confusi…

[deleted]

Re: Qwen3.6-Plus: Towards real world agents

#143
I wish these AI vendors would quit publishing comparisons with the previous generation of their competitors's models. It's just such a glaringly bad look and no one is fooled by it, even if their achievements deserve praise in their own right. The Qwen models are great and don't deserve the reputational hit that comes from dodgy marketing tactics.

Re: Qwen3.6-Plus: Towards real world agents

#146

Earlier quoted context omitted.

China has never threatened war against my country; America has. Between the two, it’s clearly safer to lean towards the Chinese options if EU ones aren’t available.

That’s incredibly naïve.

Meh, people have their own interests and values. And you can't force people to spend money no matter how much you may disagree with them

Bring on the Chinese, fuck the Americans.

Re: Qwen3.6-Plus: Towards real world agents

#147

This is their hosted-only model, not an open weight model like they’ve become known for. They got a lot of good publicity for their open weight model releases, which was the goal. The hard part is pivoting from an open weight provider to being considered as a competitor to Claude and ChatGPT. Initial reactions are mostly anger from everyone who didn’t realize that the play along was to give away the smaller models as…

4.5 is better than 4.6 though in practice. 4.6 was purely a cost savings change with enough benchmark gamification to look better.

Exactly. 3.6 plus in the exact same coding agent harness is notably worse in all of my testing compared to 3.5 plus.

The former gets stuck in ridiculous thought loops on the exact same tasks I’m testing. Fascinating really, I expected more for some reason.

Re: Qwen3.6-Plus: Towards real world agents

#148
post #17

Earlier quoted context omitted.

> I think there is a moderately large market for models like this that aren’t quite SOTA level but can be served up much cheaper. There isn't, pretty much everyone wants the best of the best.

Ever hit your daily limit on Claude Code and saw how expensive it is to pay per token?

All the time now… it’s wild how little usage you get with Opus on the Pro sub now haha

Re: Qwen3.6-Plus: Towards real world agents

#149
post #48
post #17

Earlier quoted context omitted.

> I think there is a moderately large market for models like this that aren’t quite SOTA level but can be served up much cheaper. There isn't, pretty much everyone wants the best of the best.

Nope. I get very good results from GLM 5 and 5.1. I’m not working on anything so complex and groundbreaking that I need the best. Coding is a rung on the ladder of model capability. Frontier models will grow to take on more capabilities, while smaller more focused models start becoming the economical choice for coding

GLM-5 is surprisingly good to be fair. Punches well above its weight IMO

Re: Qwen3.6-Plus: Towards real world agents

#150

Earlier quoted context omitted.

China has never threatened war against my country; America has. Between the two, it’s clearly safer to lean towards the Chinese options if EU ones aren’t available.

That’s incredibly naïve.

More naive than blithely blowing off threats of war?
Post reply on HN