Live data from Hacker News

GLM 4.5 with Claude Code

docs.z.ai

11–20 of 90 posts

Re: GLM 4.5 with Claude Code

#11

Available on OpenRouter as well for those who want to test it: https://openrouter.ai/z-ai/glm-4.5 I would be interested to know where the claim of the “killer combination” comes from. I would also like to know who the people behind Z.ai are — I haven’t heard of them before. Their plans seem crazy cheap compared to Anthropic, especially if their models actually perform better than Opus.

Actually Z.ai is a spinoff of Tsinghua University and one of the first China labs open sourcing its own large models (GLM released in 2021) . https://github.com/THUDM/GLM

It's a spinoff of the whole university?

Re: GLM 4.5 with Claude Code

#12
Hmm with the lower context length I'm wonder how it holds up for problems requiring slightly larger context given we know most models tend to degrade fairly quickly with context length.

Maybe it's best for shorter tasks or condensed context?

I find it interesting the number of models latching onto Claude codes harness. I'm still using Cursor for work and personal but tried out open code and Claude for a bit. I just miss having the checkpoints and whatnot.

Re: GLM 4.5 with Claude Code

#13
I've been using GLM 4.5 and GLM 4.5 Air for a while now. The Air model is light enough to run on a macbook pro and is useful for Cline. I can run the full GLM model on my Mac Studio, but the TPS is so slow that it's only useful for chatting. So I hooked up with openrouter to try but didn't have the same success. Any of the open weight models I try with open router give sub standard results. I get better results from Qwen 3 coder 30b a3b locally than I get from Qwen 3 Coder 480b through open router.

I'm really concerned that some of the providers are using quantized versions of the models so they can run more models per card and larger batches of inference.

Re: GLM 4.5 with Claude Code

#14
post #13

I've been using GLM 4.5 and GLM 4.5 Air for a while now. The Air model is light enough to run on a macbook pro and is useful for Cline. I can run the full GLM model on my Mac Studio, but the TPS is so slow that it's only useful for chatting. So I hooked up with openrouter to try but didn't have the same success. Any of the open weight models I try with open router give sub standard results. I get better results from…

yeah I too have heard similar concerns with Open models on OpenRouter, but haven't been able to verify it, as I don't use that a lot

Re: GLM 4.5 with Claude Code

#15
post #4

Okay, I'm going to try it, but why didn't you link the information on how to integrate it with Claude Code: https://docs.z.ai/scenario-example/develop-tools/claude Chinese software always has such a design language: - prepaid and then use credit to subscribe - strange serif font - that slider thing for captcha But I'm going to try it out now.

Ahh bugger I pasted the wrong link I had this one open in another tab..

Re: GLM 4.5 with Claude Code

#16

Available on OpenRouter as well for those who want to test it: https://openrouter.ai/z-ai/glm-4.5 I would be interested to know where the claim of the “killer combination” comes from. I would also like to know who the people behind Z.ai are — I haven’t heard of them before. Their plans seem crazy cheap compared to Anthropic, especially if their models actually perform better than Opus.

Well I'd call them the poor person's claude code, wouldnt compare it with Opus but very close to Sonnet and Kimi

Re: GLM 4.5 with Claude Code

#17

Hmm with the lower context length I'm wonder how it holds up for problems requiring slightly larger context given we know most models tend to degrade fairly quickly with context length. Maybe it's best for shorter tasks or condensed context? I find it interesting the number of models latching onto Claude codes harness. I'm still using Cursor for work and personal but tried out open code and Claude for a bit. I just m…

https://fiction.live/stories/Fiction-liveBench-Feb-21-2025/o...

Re: GLM 4.5 with Claude Code

#18
post #9

I wonder how you justify this editorialized title, and if HN mods share your justification. The linked article has no the word "killer" in it. I think this is why many people have concerns about AI. This group can't express neutral ideas. They have to hype about a simple official documentation page.

feedback accepted got rid of the killer bits

Re: GLM 4.5 with Claude Code

#19
post #7

Available on OpenRouter as well for those who want to test it: https://openrouter.ai/z-ai/glm-4.5 I would be interested to know where the claim of the “killer combination” comes from. I would also like to know who the people behind Z.ai are — I haven’t heard of them before. Their plans seem crazy cheap compared to Anthropic, especially if their models actually perform better than Opus.

> I would also like to know who the people behind Z.ai are — I haven’t heard of them before. To be clear, Z.ai are the people who built GLM 4.5, so they're talking up their own product. But to be fair, GLM 4.5 and GLM 4.5 Air are genuinely good coding models. GLM 4.5 Air costs about 10% of what Claude Sonnet does (when hosted on DeepInfra, at least), and it can perform simple coding tasks quite quickly. I haven't tes…

For agentic coding I found the price difference more modest due to prompt caching, which most GLM providers on Openrouter don't offer, but Anthropic does. Look at the cache read/write columns: https://openrouter.ai/z-ai/glm-4.5

Re: GLM 4.5 with Claude Code

#20

Earlier quoted context omitted.

Actually Z.ai is a spinoff of Tsinghua University and one of the first China labs open sourcing its own large models (GLM released in 2021) . https://github.com/THUDM/GLM

It's a spinoff of the whole university?

With a little search you can find it's a laboratory within the CS department of THU. It's a fairly large lab though, not those led by just one or two professors.
Post reply on HN