Live data from Hacker News

Qwen3.6-Plus: Towards real world agents

qwen.ai

101–110 of 235 posts

Re: Qwen3.6-Plus: Towards real world agents

#101
post #66

Pretty solid Pelican: https://gist.github.com/simonw/ca081b679734bc0e5997a43d29fad... I used the https://modelstudio.alibabacloud.com/ API to generate that one, which required signing up for an account and attaching PayPal billing - but it looks like OpenRouter are offering it for free right now so I could have used that: https://openrouter.ai/qwen/qwen3.6-plus:free

Pelican is drafting rear peloton

Re: Qwen3.6-Plus: Towards real world agents

#102
post #59
post #17

Earlier quoted context omitted.

> I think there is a moderately large market for models like this that aren’t quite SOTA level but can be served up much cheaper. There isn't, pretty much everyone wants the best of the best.

No. Right now I'm upset that Google has removed (or at least is in the process of removing) the Gemini 2.0 flash model. We use it for some pretty basic functionality because it's cheap and fast and honestly good enough for what we use it for in that part of our app. We're being forced to "upgrade" to models that are at least 2.5 times as expensive, are slower and, while I'm sure they're better for complex tasks, don'…

In case you don't know, Gemini 2.5 flash is hosted on DeepInfra. They also have 1.5 flash but not 2.0 flash.

I have no affiliation with DeepInfra. I use them, because they host open-source models that are good.

Re: Qwen3.6-Plus: Towards real world agents

#103
post #66

Pretty solid Pelican: https://gist.github.com/simonw/ca081b679734bc0e5997a43d29fad... I used the https://modelstudio.alibabacloud.com/ API to generate that one, which required signing up for an account and attaching PayPal billing - but it looks like OpenRouter are offering it for free right now so I could have used that: https://openrouter.ai/qwen/qwen3.6-plus:free

they're going to start training a pelican riding a bike specifically on these models soon. it's the key global benchmark!

Re: Qwen3.6-Plus: Towards real world agents

#104
post #76

Earlier quoted context omitted.

> Initial reactions are mostly anger from everyone who didn’t realize that the play along was to give away the smaller models as advertising, not because they were feeling generous. The naivety around this has been staggering quite frankly. All of a sudden, people thinking that meta etc are releasing free models because they believe in open access and distribution of knowledge. No, they just suck comparatively. There…

I don't think there's so much naivety. People can be aware of the the plan and still be frustrated and disappointed when it happens.

For a brief moment there were a lot of comments about how Chinese tech companies are our saviors in the age of AI because they were releasing their models. It was an edgy contrarian take that was getting a lot of traction, mostly from commenters who were unfamiliar with Alibaba and thought it was the anti-Big-tech

Re: Qwen3.6-Plus: Towards real world agents

#105
post #38

Earlier quoted context omitted.

so if China has the data good, us has the data bad, got it lol. us actually has laws around this and they arent sharing very much with thr us gov today. china shares 100% as required by law. and neither care much about "how long do i cook eggs for", but they do care about code generation a lot.

From an espionage perspective your own government is the safest. But from a civil rights perspective your own government is your most immediate threat. China isn't going to arrest me for my opinions on Netanyahu, my own government could And the US government has repeatedly shown that it is very interested in collecting all the data available, just like China. In China this is simply done in the open while the US has…

> China isn't going to arrest me for my opinions on Netanyahu, my own government could

I don't know whether you really believe this or it was an off the cuff remark. China is not going to tell you why they plan to arrest you. China is not a benevolent dictatorship.

Re: Qwen3.6-Plus: Towards real world agents

#106

This is their hosted-only model, not an open weight model like they’ve become known for. They got a lot of good publicity for their open weight model releases, which was the goal. The hard part is pivoting from an open weight provider to being considered as a competitor to Claude and ChatGPT. Initial reactions are mostly anger from everyone who didn’t realize that the play along was to give away the smaller models as…

Opus was released in Feb 2026. Even though it feels like a long 2 months has passed, its' not really clear that they were developing this as a competitor to that product. There's nothing really strange about not competing directly with the best, but rather showing whom you are as good as .

I don’t know why anyone would do the mental backflips to defend this.

They posted charts with logos for Claude and others. You had to read the fine details to realize they weren’t comparing to the latest offerings from those companies. They were counting on you not noticing.

There’s zero reason to compare to old models unless you’re trying to mislead.

Re: Qwen3.6-Plus: Towards real world agents

#107
post #96

Earlier quoted context omitted.

> not an open weight model like they’ve become known for. Right, they state that they'll release "smaller" variants openly at some point, with few details as to what that means. Will there be a ~300B variant as with Qwen 3.5? The blog post doesn't say.

I'm not interested in adopting an inferior closed source weight from a geopolitical rival. The open source weights argument was the one thing China had going and that I was seriously cheering them on for. They could have been our saviors and disrupted the US tech giants - and if it was open, I'd have welcomed it. Now they show their true colors. They want to train models on our engineering to replace us, while simult…

> I'm not interested in adopting an inferior closed weights model from a geopolitical rival.

That's a very reasonable stance. It doesn't change the fact that we do have plenty of local models (up to and including Qwen 3.5) that are still quite useful.

Re: Qwen3.6-Plus: Towards real world agents

#108
post #96

Earlier quoted context omitted.

> not an open weight model like they’ve become known for. Right, they state that they'll release "smaller" variants openly at some point, with few details as to what that means. Will there be a ~300B variant as with Qwen 3.5? The blog post doesn't say.

I'm not interested in adopting an inferior closed source weight from a geopolitical rival. The open source weights argument was the one thing China had going and that I was seriously cheering them on for. They could have been our saviors and disrupted the US tech giants - and if it was open, I'd have welcomed it. Now they show their true colors. They want to train models on our engineering to replace us, while simult…

Whereas I as a Canadian am absolutely eager to see a serious competitor from a rival to the US because sending money south to Anthropic and OpenAI who think it's ok to spy on (or worse) their non-American customers, and are headquartered in a country that is trying to crush my country's economy, interfere in our domestic politics, and put us out of work and making threats on political allies.

I'd prefer them to be open weight, but I'd love to sub a decent competitive coding plan from a European or Chinese provider. Right now they're not quite there. If closing it and charging for it brings them closer to competitive, that's ok.

If the US tech and AI industry long term wants customers and a broad market outside of their own domestic base, they need to reconsider who they are bending the knee to, and how they are defining their policies in relation to the Trump administration.

Bring on the Chinese competition.

Re: Qwen3.6-Plus: Towards real world agents

#109

I'll diverge from some of these comments, I don't find it misleading to compare to Opus 4.5. I can remember how good Opus 4.5 was. If I'm considering using this, it's most informative to me to compare to the model it's closest to that I have familiarity with. I'm obviously not switching to this if I want the best model. I'm switching if I'm hopeful that the smaller versions are close to it, or if I want to have more…

Yes, honestly, Opus 4.6 and GPT 5.4 were mostly not really noticeable improvements over 4.5 and 5.3 respectively. If we were stuck at 4.5 levels but at 1/10th of the price, I'll take it.
Post reply on HN