Live data from Hacker News

Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

qwen.ai

231–240 of 400 posts

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#232
post #71

Ok I find it funny that people compare models and are like, opus 4.7 is SOTA and is much better etc, but I have used glm 5.1 (I assume this comes form them training on both opus and codex) for things opus couldn't do and have seen it make better code, haven't tried the qwen max series but I have seen the local 122b model do smarter more correct things based on docs than opus so yes benchmarks are one thing but realit…

I wonder why glm is viewed so positively. Every time I try to build something with it, the output is worse than other models I use (Gemini, Claude), it takes longer to reach an answer and plenty of times it gets stuck in a loop.

I think it offers a very good tradeoff of cost vs competency

4.7 is better, but its also wildly expensive

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#233
post #166

Earlier quoted context omitted.

The Western Zen? In my experience it is downgraded from being a religion to being a system of practice which relieves it of the broader Mahayana cosmology. But I would suggest the dogma is less obvious but still there, often just somewhere else, such as in its own limitations, or in a philosophical container at a higher level such as scientism.

All Zen is about releasing those attachments. Granted it's pretty hard, because if you succeed, you're enlightened. East, West, Religion, Practice… From a Zen perspective, you're just troubling your mind with binaries and conflict.

Ah and there is the dogma -- the otherness of the enlightened.

The binaries still functionally exist. I see a lot of value in reflective practices. At the same time it seems unlikely to me that the point of existing is to not trouble your mind.

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#234

Earlier quoted context omitted.

> I've used Claude for many months now. Since February I see a stark decline in the work I do with it. I find myself repeating the following pattern: I use an AI model to assist me with work, and after some time, I notice the quality doesn't justify the time investment. I decide to try a similar task with another provider. I try a few more tests, then decide to switch over for full time work, and it feels like it's a…

I wonder about this. I see two obvious possibilities (if we ignore bias): 1. The models are purposefully nerfed, before the release of the next model, similar to how Apple allegedly nerfed their older phones when the next model was out. 2. You are relying more and more on the models and are using your talent less and less. What you are observing is the ratio of your vs. the model’s work leaning more and more to the m…

I don’t think the providers intentionally nerf the models to make the new one look better. It’s a matter of them being stingy with infrastructure, either by choice to increase profit and/or sheer lack of resources to keep n+1 models deployed in parallel without deprecating older ones when a new one is released.

I’d prefer providers to simply deprecate stuff faster, but then that would break other people’s existing workflows.

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#235
post #214

Earlier quoted context omitted.

> The big kicker for GLM for me is I can use it in Pi, or whatever harness I like. Yes, but... isn't the same true for Opus and all the other models too?

Opus is about 7 times more expensive than GLM with API pricing. And since you can only use the Opus subscription plan in CC, you're essentially locked into API pricing for Pi and any other harness. So you're either paying $1000's for Opus in Pi, or $30/month for GLM in Pi. If the results are mostly equivalent that's an easy choice for most of us.

Perhaps I'm being extremely daft: If the API is 7 times more expensive, then why is it $1000 vs $30? Or is there a GLM subscription one can use with Pi? Certainly not available in my (arguably outdated) Pi.

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#236

Earlier quoted context omitted.

> I've tried using qwen and deepseek but they can't even output documents What agent harness did you use? Usually, "write_file", "shell_exec" or similar is two of the first tools you add to an agent harness, after read_file/list_files. If it doesn't have those tools, unsure if you could even call it a agent harness in the first place.

Sorry for the confusion, I was actually talking about their Web based chat. Since most of my work is governance and docs, I just use their Web chats and they just refuse to output proper documents like Claude or Chatgpt do.

You're not giving an AI command line access to your work computer? How do you expect to keep up? /s

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#238
post #235

Earlier quoted context omitted.

Opus is about 7 times more expensive than GLM with API pricing. And since you can only use the Opus subscription plan in CC, you're essentially locked into API pricing for Pi and any other harness. So you're either paying $1000's for Opus in Pi, or $30/month for GLM in Pi. If the results are mostly equivalent that's an easy choice for most of us.

Perhaps I'm being extremely daft: If the API is 7 times more expensive, then why is it $1000 vs $30? Or is there a GLM subscription one can use with Pi? Certainly not available in my (arguably outdated) Pi.

I'm not the OP, but it's the latter. I'm currently using the "Lite" GLM subscription with OpenCode, for example. I'm not using it very heavily, but I haven't come close to hitting the limits, whereas I burned through my weekly limits with Claude very regularly.

Re: Qwen3.6-Max-Preview: Smarter, Sharper, Still Evolving

#239

Earlier quoted context omitted.

Sorry for the confusion, I was actually talking about their Web based chat. Since most of my work is governance and docs, I just use their Web chats and they just refuse to output proper documents like Claude or Chatgpt do.

You're not giving an AI command line access to your work computer? How do you expect to keep up? /s

You give it command line access in a VM...
Post reply on HN