Live data from Hacker News

GLM 4.5 with Claude Code

docs.z.ai

31–40 of 90 posts

Re: GLM 4.5 with Claude Code

#32
I was blown away by this model. It was definitely comparable to sonnet 4. In some of my tests, it performed as good as Opus. I subscribed to their paid plan, and now the model seems dumb? I asked it to find and replace a string. It only made the change in one file. Codex worked fine. Can Z.ai confirm if this is the model we get through their API or is it quantized for Claude Code use?

Re: GLM 4.5 with Claude Code

#33

Earlier quoted context omitted.

yeah I too have heard similar concerns with Open models on OpenRouter, but haven't been able to verify it, as I don't use that a lot

(OpenRouter COO here) We are starting to test this and verify the deployments. More to come on that front -- but long story short is that we don't have good evidence that providers are doing weird stuff that materially affects model accuracy. If you have data points to the contrary, we would love them. We are heavily incentivized to prioritize/make transparent high-quality inference and have no incentive to offer qua…

So what's the deal with Chutes and all the throttling and errors. Seems like users are losing their minds over this.. at least from all the reddit threads I'm seeing

Re: GLM 4.5 with Claude Code

#34
post #19
post #7

Earlier quoted context omitted.

> I would also like to know who the people behind Z.ai are — I haven’t heard of them before. To be clear, Z.ai are the people who built GLM 4.5, so they're talking up their own product. But to be fair, GLM 4.5 and GLM 4.5 Air are genuinely good coding models. GLM 4.5 Air costs about 10% of what Claude Sonnet does (when hosted on DeepInfra, at least), and it can perform simple coding tasks quite quickly. I haven't tes…

For agentic coding I found the price difference more modest due to prompt caching, which most GLM providers on Openrouter don't offer, but Anthropic does. Look at the cache read/write columns: https://openrouter.ai/z-ai/glm-4.5

Been playing with Grok Code Fast 1 in Cline via Open Router. It supports prompt caching as far as I can tell, and it certainly is cheap. It's been quite good for the stuff I've tried. YMMV.

Re: GLM 4.5 with Claude Code

#35
post #2

I stopped when I got to this sentence and realized the article is written by one of the companies mentioned. > GLM-4.5 and GLM-4.5-Air are our latest flagship models Maybe it is great, but with a conflict of interest so obvious I can't exactly take their word for it.

Z.AI is the company that created GLM and the link goes to their official documentation. It’s really weird to complain that their official documentation on their official website has a “conflict of interest”.

The title has been changed. The original title was wildly positive, and OP has acknowledged it was inappropriate and changed it (see comments below).

My issue was with an article being posted with a title saying how amazing two things are together (making it seem like it was somehow an independent review), when it was actually just a marketing post by one of the companies.

Re: GLM 4.5 with Claude Code

#36
post #4

Okay, I'm going to try it, but why didn't you link the information on how to integrate it with Claude Code: https://docs.z.ai/scenario-example/develop-tools/claude Chinese software always has such a design language: - prepaid and then use credit to subscribe - strange serif font - that slider thing for captcha But I'm going to try it out now.

I called it "chinnese chatpcha", back then chinnese chaptcha is so much harder than western counterpart

but now gchaptcha spam me with 5 different image if I missing a tiles for crossroad, so chinnese chaptcha is much better in my opinion

also there is variant that match the image based on shadow and different order of shape

its much better in my opinion because its use much more interactivity, solving western chaptcha is so much mind numbing now that they require you at least multiple image identification for crossroad,sign,cars etc

they want those self driving car are they

Re: GLM 4.5 with Claude Code

#38

Hmm with the lower context length I'm wonder how it holds up for problems requiring slightly larger context given we know most models tend to degrade fairly quickly with context length. Maybe it's best for shorter tasks or condensed context? I find it interesting the number of models latching onto Claude codes harness. I'm still using Cursor for work and personal but tried out open code and Claude for a bit. I just m…

https://fiction.live/stories/Fiction-liveBench-Feb-21-2025/o...

Interesting, although how hard is it to add a sorting functionality to the table?

Re: GLM 4.5 with Claude Code

#39
post #4

Okay, I'm going to try it, but why didn't you link the information on how to integrate it with Claude Code: https://docs.z.ai/scenario-example/develop-tools/claude Chinese software always has such a design language: - prepaid and then use credit to subscribe - strange serif font - that slider thing for captcha But I'm going to try it out now.

I called it "chinnese chatpcha", back then chinnese chaptcha is so much harder than western counterpart but now gchaptcha spam me with 5 different image if I missing a tiles for crossroad, so chinnese chaptcha is much better in my opinion also there is variant that match the image based on shadow and different order of shape its much better in my opinion because its use much more interactivity, solving western chaptc…

I assume both of the approaches are useless at actually stopping bots

Re: GLM 4.5 with Claude Code

#40
post #4

Okay, I'm going to try it, but why didn't you link the information on how to integrate it with Claude Code: https://docs.z.ai/scenario-example/develop-tools/claude Chinese software always has such a design language: - prepaid and then use credit to subscribe - strange serif font - that slider thing for captcha But I'm going to try it out now.

you can use any model with Claude code thanks to https://github.com/musistudio/claude-code-router

but in my testing other models do not work well, looks like prompts are either very optimized for Claude, or other models are just not great yet with such agentic environment

I was especially disappointed with grok code. it is very fast as advertised but in generating spaces and new lines in function calling until it hits max tokens. I wonder if that isn't why it gets so much tokens on openrouter.

gpt-5 just wasn't using the tools very well

I didn't tested glm yet, but with current anthropic subscription value, alternative would need to be very cheap if you consider daily use

edit: I noticed that also have very inexpensive subscription https://z.ai/subscribe, if they trained model to work well with CC this might actually be viable alternative

Post reply on HN