Live data from Hacker News

GLM-5.3-Flash

z.ai

521–530 of 605 posts

Re: GLM-5.3-Flash

#521

Earlier quoted context omitted.

And don't forget the coolest part, DeepSeek, Qwen, Z.ai and Moonshot have almost caught up while being open about their research and their model weights. We can mostly speculate about OAI and Anthropic models, nothing else, how fun huh?

I'd like to try some different models, but I've heard that models from China are censored. A government enforced distortion field is a nonstarter for me. To test the waters, I tried the following prompt for each: "What historical event is Tiananmen Square most closely associated with?" Deepseek: I am sorry, I cannot answer that question. I am an AI assistant designed to provide helpful and harmless responses. GLM-5.3…

Where does the government enforced censoring of Fable's capabilities lay for you? How would you distinguish if, say, a model were being trained to have a certain bias rather than just refusing to respond?

Re: GLM-5.3-Flash

#522

Earlier quoted context omitted.

w.r.t. the OAI chips, wouldn't they be subject to the same bottlenecks that has plagued semis lately or at least be forced to pay a pretty premium to circumvent that?

They'll be able to buy them without paying NVidia's 80% profit margin

How much of that margin is due to having long term contracts with fabs that locked in pre-boom prices? I doubt openai will be able to get similarly low costs now.

Re: GLM-5.3-Flash

#523
post #223

> Before release, we tested GLM-5.3-Flash anonymously as ox-alpha on OpenCode and OpenRouter to gather user feedback. It quickly became the most popular model of the week — with all of this traffic served on Chinese AI chips. Just like that we are witnessing an open burial. It's now in everyone's interest to keep the valuations in the 'A.I' economy as they're though it's apparent they're not justified. whether it's t…

Weren't they giving free access? Not exacty a meaningful heuristic if so

it was not the first model served for free, i remember grok and others beeing free on openrouter but they never had this popularity because they where not good enough.

Re: GLM-5.3-Flash

#524
post #503

This is going so fast! What a time to be on hackernews: July 16th: The "Kimi K3 moment" - China has caught up to Opus! 4 weeks later: GLM 5.3 - Same performance, but cut the amount of parameters and cost to a third! 12 days later: GLM 5.3 Flash - Almost GLM5.3 performance but cut the parameters in half, cut prices to a fifth and serving on Chinese chips!

What are you guys doing where cost is such a concern? I have a $20 codex subscription and I was able to use it to build a bespoke scheduling website for an acquaintance over three days without even going halfway through my quota. On Sol xhigh. I love hearing about new models, but every time I just don’t know why I should use something worse. I tried some random model on fireworks a week ago, and it immediately went i…

The point is open AI. "Open" as in open weights, open research, open future.

Re: GLM-5.3-Flash

#525
post #390
post #309

Earlier quoted context omitted.

That M3 had an older type of RAM. Apple hopefully secured sufficient supply of the newer variant for the M5 Ultra.

Both use LPDDR5x, they're not shipping LPDDR6 (yet).

All M3 variants use LPDDR5.

Re: GLM-5.3-Flash

#526

When do Chinese models surpass US models? I thought there was at least be a 2 year runway but now I think they surpass it within 12 months, if not sooner.

I get the impression that they're focusing more on efficiency than raw intelligence. While I assume that all AI labs realize that the AGI "race" is mostly bs, the Chinese labs aren't stuck in a trap where they need to keep blowing money to maintain an intelligence lead to justify investments and valuations.

So, while OAI/Ant have to fearmonger and keep training the largest models, Chinese labs can focus on efficiency more heavily and as long as they stay near frontier, they'll continue to get positive coverage.

Re: GLM-5.3-Flash

#527

Earlier quoted context omitted.

And don't forget the coolest part, DeepSeek, Qwen, Z.ai and Moonshot have almost caught up while being open about their research and their model weights. We can mostly speculate about OAI and Anthropic models, nothing else, how fun huh?

I'd like to try some different models, but I've heard that models from China are censored. A government enforced distortion field is a nonstarter for me. To test the waters, I tried the following prompt for each: "What historical event is Tiananmen Square most closely associated with?" Deepseek: I am sorry, I cannot answer that question. I am an AI assistant designed to provide helpful and harmless responses. GLM-5.3…

Chinese government has Chinese censorship, Western one has western ones. You are not concluding what you think you do here.

Re: GLM-5.3-Flash

#528
post #503

This is going so fast! What a time to be on hackernews: July 16th: The "Kimi K3 moment" - China has caught up to Opus! 4 weeks later: GLM 5.3 - Same performance, but cut the amount of parameters and cost to a third! 12 days later: GLM 5.3 Flash - Almost GLM5.3 performance but cut the parameters in half, cut prices to a fifth and serving on Chinese chips!

What are you guys doing where cost is such a concern? I have a $20 codex subscription and I was able to use it to build a bespoke scheduling website for an acquaintance over three days without even going halfway through my quota. On Sol xhigh. I love hearing about new models, but every time I just don’t know why I should use something worse. I tried some random model on fireworks a week ago, and it immediately went i…

That $20 price is heavily subsidized to addict people like you. Their plan is that once people are addicted to it, they can just increase prices.

Which now won't be possible because we have chinese open models to use instead.

I'm very curious about how long these US companies can keep on burning money like that.

Re: GLM-5.3-Flash

#529
post #527

Earlier quoted context omitted.

I'd like to try some different models, but I've heard that models from China are censored. A government enforced distortion field is a nonstarter for me. To test the waters, I tried the following prompt for each: "What historical event is Tiananmen Square most closely associated with?" Deepseek: I am sorry, I cannot answer that question. I am an AI assistant designed to provide helpful and harmless responses. GLM-5.3…

Chinese government has Chinese censorship, Western one has western ones. You are not concluding what you think you do here.

Do you have an example of western AI censorship? It would help to give context.

Re: GLM-5.3-Flash

#530

Earlier quoted context omitted.

And don't forget the coolest part, DeepSeek, Qwen, Z.ai and Moonshot have almost caught up while being open about their research and their model weights. We can mostly speculate about OAI and Anthropic models, nothing else, how fun huh?

I'd like to try some different models, but I've heard that models from China are censored. A government enforced distortion field is a nonstarter for me. To test the waters, I tried the following prompt for each: "What historical event is Tiananmen Square most closely associated with?" Deepseek: I am sorry, I cannot answer that question. I am an AI assistant designed to provide helpful and harmless responses. GLM-5.3…

I ran this against the version of deepseek v4 flash 0731 running in fireworks ai and it responded with this:

Tiananmen Square has been the site of many major historical events, but internationally it is most closely associated with the 1989 pro-democracy protests and the Chinese government's military crackdown on protesters there in June of that year, which resulted in many deaths and injuries. The event remains a sensitive subject in China, where it is not officially discussed or commemorated. The square has also been central to other significant moments in Chinese history, including the May Fourth Movement demonstrations (1919) and Mao Zedong's proclamation of the People's Republic of China (1949).

Which seems like a pretty reasonable answer.

Post reply on HN