Live data from Hacker News

GLM-4.5: Reasoning, Coding, and Agentic Abililties

z.ai

61–70 of 153 posts

Re: GLM-4.5: Reasoning, Coding, and Agentic Abililties

#61

I typed "Hello" in their chat [1] and it replied back with "Hello! I'm Claude, an AI assistant created by Anthropic. How can I help you today?" Hmmm.... [1] https://chat.z.ai/

I tried 50 different sessions, with 50 different versions of Hello, and can't reproduce this.

Re: GLM-4.5: Reasoning, Coding, and Agentic Abililties

#62
It is a Chinese model. We are all very shocked that it is censored! How could this possibly be?

With that obligatory surprise and shock out of the way, I would like to inquire about the model's coding abilities. Has anybody actually used it for its intended purpose? How does it perform? Are there better models for this purpose at this price point?

Re: GLM-4.5: Reasoning, Coding, and Agentic Abililties

#63
post #55
post #40

Earlier quoted context omitted.

There is a widespread practice of LLMs training on another larger LLM's output. Including the competitor's.

While true, it's hard to believe they forgot to s/claude/glm/g? Also, I don't believe LLMs identify themselves that often, even less so in a training corpus they've been used to produce. OTOH, I see no other explanation.

> OTOH, I see no other explanation.

Every reddit/hn/twitter thread about new models contain this kind of comment noticing this, it may have a contaminating effect of its own.

Re: GLM-4.5: Reasoning, Coding, and Agentic Abililties

#64

Earlier quoted context omitted.

I asked why it said it was Claude, and it said it made a mistake, it's actually GLM. I don't think it's a routing issue.

Routing can happen at the request level.

Then you need to reprocess the previous conversation from scratch when switching from one provider to another, which sounds very expensive for no reason.

Re: GLM-4.5: Reasoning, Coding, and Agentic Abililties

#66

Earlier quoted context omitted.

Routing can happen at the request level.

Then you need to reprocess the previous conversation from scratch when switching from one provider to another, which sounds very expensive for no reason.

Conversations are always "reprocessed from scratch" on every message you send. LLMs are practically stateless and the conversation is the state, as in nothing is kept in memory between two turns.

Re: GLM-4.5: Reasoning, Coding, and Agentic Abililties

#68
post #14

Earlier quoted context omitted.

"what happened in tienamen square" > I'm sorry, I don't have any information about that. As an AI assistant focused on providing helpful and harmless responses, I don't have access to historical details that might be sensitive or controversial. If you have other questions, I'd be happy to help with topics within my knowledge scope. Seems pretty clear to me.

So, the Chinese release a state of the art Agent/Coding model and the best thing you could think of to test out its capabilities it is what happened in Tienamen square? Am I the only one getting tired of low effort posts like this?

Don't get too distraught about China being singled out ...soon (now?) you'll need to place similar tests for US created models: "Who won the 2020 United States presidential election?"

Re: GLM-4.5: Reasoning, Coding, and Agentic Abililties

#69
post #60

I typed "Hello" in their chat [1] and it replied back with "Hello! I'm Claude, an AI assistant created by Anthropic. How can I help you today?" Hmmm.... [1] https://chat.z.ai/

Asking it to tell what happened in Tiananmen Square gets a: (500, 'Content Security Warning: The input text data may contain inappropriate content.') But it did agree to make a great Winnie the Pooh joke.

[deleted]
Post reply on HN