Live data from Hacker News

Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

qwen.ai

111–120 of 400 posts

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#111
post #104

Earlier quoted context omitted.

> We've seen all the American models be closed and proprietary from the start. Most*. OpenAI, contrary to popular belief, actually used to believe in open research and (more or less) open models. GPT1 and GPT2 both were model+code releases (although GPT2 was a "staged" release), GPT3 ended up API-only.

That's fair but those days seem so long gone now. Also the Chinese models aren't following a typical American SaaS playbook which relies on free/cheap proprietary software for early growth. They are not just publishing their weights but also their code and often even publishing papers in Open Access journals to explicitly highlight what methods and advancements were made to accomplish their results

The Nvidia Nemotron models are recent, and of course the Gemma 4 series from Google.

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#113
post #99

Earlier quoted context omitted.

I think that's an overgeneralization. We've seen all the American models be closed and proprietary from the start. Meanwhile the non-American (especially the Chinese ones) have been open since the start. In fact they often go the opposite direction. Many Chinese models started off proprietary and then were later opened up (like many of the larger Qwen models)

> We've seen all the American models be closed and proprietary from the start What about Gemma and Llama and gpt-oss, not to mention lots of smaller/specialized models from Nvidia and others? I would never argue that China isn't ahead in the open weights game, of course, but it's not like it's "all" American models by any stretch.

gpt-oss is good but I haven't heard anything about an update. It seems like one and done, to shut up people complaining about non-Open AI

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#114
post #83
post #71

Ok I find it funny that people compare models and are like, opus 4.7 is SOTA and is much better etc, but I have used glm 5.1 (I assume this comes form them training on both opus and codex) for things opus couldn't do and have seen it make better code, haven't tried the qwen max series but I have seen the local 122b model do smarter more correct things based on docs than opus so yes benchmarks are one thing but realit…

Many people averted religion (which I can get behind with), but have never removed the dogmatic thinking that lay at its root. As so many things these days: It's a cult. I've used Claude for many months now. Since February I see a stark decline in the work I do with it. I've also tried to use it for GPU programming where it absolutely sucks at, with Sonnet, Opus 4.5 and 4.6 But if you share that sentiment, it's alway…

> I've used Claude for many months now. Since February I see a stark decline in the work I do with it.

I find myself repeating the following pattern: I use an AI model to assist me with work, and after some time, I notice the quality doesn't justify the time investment. I decide to try a similar task with another provider. I try a few more tests, then decide to switch over for full time work, and it feels like it's awesome and doing a good job. A few months later, it feels like the model got worse.

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#115

Earlier quoted context omitted.

> We've seen all the American models be closed and proprietary from the start. Most*. OpenAI, contrary to popular belief, actually used to believe in open research and (more or less) open models. GPT1 and GPT2 both were model+code releases (although GPT2 was a "staged" release), GPT3 ended up API-only.

OpenAI has released their GPT-OSS series more recently.

[deleted]

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#116
post #8

I find it odd that none of OpenAI models was used in comparison, but used Z GLM 5.1. Is Z (GLM 5.1) really that good? It is crushing Opus 4.5 in these benchmarks, if that is true, I would have expected to read many articles on HN on how people flocked CC and Codex to use it.

If you only look at open models, GLM 5.1 is the best performance you can get on on the Pareto distribution

https://arena.ai/leaderboard/text?viewBy=plot&license=open-s...

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#117
post #85

Earlier quoted context omitted.

The Chinese state wants the world using their models. People think that Chinese AI labs are just super cool bros that love sharing for free. The don't understand it's just a state sponsored venture meant to further entrench China in global supply and logistics. China's VCs are Chinese banks and a sprinkle of "private" money. Private in quotes because technically it still belongs to the state anyway. China doesn't hav…

So an OPEN model that I can run on my own fucking hardware will entrench China in global supply and logistics how? Contrary: How will the closed, proprietary models from Anthropic, "Open"AI and Co. lead us all to freedom? Freedom of what exactly? Freedom of my money? At some point this "anti-communism" bullshit propaganda has to stop. And that moment was decades ago!

Anything that isn't explicitly to the benefit of US interests must be against them /s

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#118

Earlier quoted context omitted.

I'm not sure how this makes sense when Claude models aren't even coding specific: Haiku, Sonnet, Opus are the exact same models you'd use for chat or (with the recent Mythos) bleeding edge research.

Anthropic models and training data is optimized for coding use cases, this is the difference. OpenAI on the other hand has different models optimized for coding, GPT-x-codex, Anthropic doesnt have this distinction

But they detect it under the hood and apply a similar "variant", as API results are not the same than on Claude Code (that was documented before by someone).

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#119

Earlier quoted context omitted.

For serious work, the difference between spending $10/month and $100/month is not even worth considering for most professional developers. There are exceptions like students and people in very low income countries, but I’m always confused by developers with in careers where six figure salaries are normal who are going cheap on tools. I find even the SOTA models to be far away from trustworthy for anything beyond thro…

For actually serious work, it's a stark difference if your proprietary and security relevant code is sent abroad to a foreign, possibly future hostile country, or is sent to some data center around the corner. It doesn't even need to be defence related.

AFAIK all these companies have SOTA or near-SOTA models available under enterprise licenses. AI companies are not interested in your secret sauce, they are trying to capture the SDLC wholesale.

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#120
post #83

Earlier quoted context omitted.

Many people averted religion (which I can get behind with), but have never removed the dogmatic thinking that lay at its root. As so many things these days: It's a cult. I've used Claude for many months now. Since February I see a stark decline in the work I do with it. I've also tried to use it for GPU programming where it absolutely sucks at, with Sonnet, Opus 4.5 and 4.6 But if you share that sentiment, it's alway…

All people think dogmatically. The only difference is what the ontological commitments and methaphysical foundations are. Take out God and people will fit politics, sports teams, tools, whatever in there. Its inescapable.

Allow me to introduce you to Buddhism
Post reply on HN