Live data from Hacker News

GLM-5.3-Flash

z.ai

341–350 of 605 posts

Re: GLM-5.3-Flash

#341

You guys read Z.ai's terms of service, right? Broad and perpetual license over inputs and outputs, and even your name and profile picture. Vague prohibitions on whatever may harm Z.ai’s "interests" or even the "national interests" of any country. Vague prohibitions on "disturbing" or "inappropriate" content, whatever that is. Vague prohibitions on discussing Z.ai, even my posting this comment violates it. Can ban you…

As long as the weights are open, who cares? You can rely on someone else for inference.

Re: GLM-5.3-Flash

#342
post #300

Earlier quoted context omitted.

Trolling. GLM is heavily distilled from Gemini.

Source? GLM is great for coding and Gemini is barely useful in coding, to be generous. I highly suspect that the Gemini Google uses internally is very different from what they offer in Antigravity.

I can't speak for GLM as I haven't tested it much, but my experience with running DeepSeek V4 locally is the first time I prompted it with "Explain your capabilities to me", it responded that it was Gemini, a multi-modal model. I've seen others see the same with DSv4, as well as the "hallucinated" multi-modal nature. I would not be surprised if GLM similarly was partially (heavily?) distilled off of Gemini.

A fun test would be to compare the logits for "gemini", "claude", etc for a continuation of "I am " on all these models. I'm sure that e.g. GLM, Qwen, DS are dominant, but I'd be curious to see the next highest contenders, and how they compare to each other.

Re: GLM-5.3-Flash

#343
post #300

Earlier quoted context omitted.

Trolling. GLM is heavily distilled from Gemini.

Source? GLM is great for coding and Gemini is barely useful in coding, to be generous. I highly suspect that the Gemini Google uses internally is very different from what they offer in Antigravity.

No, it's the same internally and externally.

Gemini 3.7 Flash is a pretty great model IMO. You shouldn't compare it to Opus, Sol, K3, etc since it's a much smaller model but it's a little better compared to Sonnet, Luna or Terra, etc.

Re: GLM-5.3-Flash

#344
post #7

Weights on HF here: https://huggingface.co/zai-org/GLM-5.3-Flash I decided to take the plunge and get myself four sparks at a decent price (and bought the QSFP cables from AliExpress because they are literally 1/2 the price of Amazon), even knowing Apple was going to release new hardware and there's probably a spark 2 on the horizon. It looks like this is going to be a decent fit for what I need. I've been experiment…

Sparks don't have enough memory bandwidth, for the same 20k you're better off buying RTX or Apple M5 Ultra machines.

Re: GLM-5.3-Flash

#345
post #185
post #79

Chinese labs are so used to manipulating benchmarks to try to flatter inferior models that when they finally have one that's really pretty good I think the official announcement here undersells it. https://deepswe.datacurve.ai/ That's pretty solid. Smarter and cheaper than Luna xhigh, not as smart but less expensive than Luna max. Smashes deepseek v4 flash, and even worse it matches v4 pro at a tiny fraction the cost…

I don't know how anyone can actually use Luna max on ANY real workload. I've had Sol orchestrate a bunch of Luna agents, these agents were explicitly given small chunks of larger objectives and they still filled their entire context windows with just reasoning tokens, until compaction hit, and then reasoning again. I've probably wasted a good 40% of my weekly usage on Luna Max agents just thinking and not writing a s…

Are you sure you actually spun up Luna sub agents? Sol up until recently could only spin up Sol and Terra agents and would even name them "Luna" despite not being so, you can check in your usage whether it was Luna or not. I believe now it's fixed though.

https://www.reddit.com/r/codex/comments/1vj3hhn/wait_so_sol_...

https://www.reddit.com/r/codex/comments/1vp0rig/sol_can_fina...

Re: GLM-5.3-Flash

#346
post #5

> with all of this traffic served on Chinese AI chips RIP Nivida shareholders

I don't see a situation where subscription payers move outside American LLMs (chatgpt, claude, gemini) And I don't see a situation where serious API payers are OK with handing the Chinese state all their data. Like manufactures of decades past did and learned a hard, even existential, lesson for it. The state mantra has been "Collect and Copy" for a long time now, tech just hasn't had that moment to experience it yet…

Casual consumers are using American models because their usage is low. As usage scales, the economics heavily favor open weight models. The API pricing from American companies is absurd. This is particularly true in an enterprise setting.

Re: GLM-5.3-Flash

#347
post #186

This is going so fast! What a time to be on hackernews: July 16th: The "Kimi K3 moment" - China has caught up to Opus! 4 weeks later: GLM 5.3 - Same performance, but cut the amount of parameters and cost to a third! 12 days later: GLM 5.3 Flash - Almost GLM5.3 performance but cut the parameters in half, cut prices to a fifth and serving on Chinese chips!

except besides benchmarks, most of these models don't meet reliability of Sol/Opus in coding work. Opus unfortunately talks very weirdly so not a great out of the box experience

[deleted]

Re: GLM-5.3-Flash

#348
post #186

This is going so fast! What a time to be on hackernews: July 16th: The "Kimi K3 moment" - China has caught up to Opus! 4 weeks later: GLM 5.3 - Same performance, but cut the amount of parameters and cost to a third! 12 days later: GLM 5.3 Flash - Almost GLM5.3 performance but cut the parameters in half, cut prices to a fifth and serving on Chinese chips!

except besides benchmarks, most of these models don't meet reliability of Sol/Opus in coding work. Opus unfortunately talks very weirdly so not a great out of the box experience

Opus 5 is the least reliable frontier-class model in the market

Re: GLM-5.3-Flash

#349
post #191

Earlier quoted context omitted.

That’s pretty funny to say when the EU claims GDPR applies worldwide.

What they don’t do. They claim that the GDPR applies if you provide your service in the EU, and that’s a valid claim.

They claim it applies to EU citizens when both they and the service are outside the EU as well.

Re: GLM-5.3-Flash

#350

You guys read Z.ai's terms of service, right? Broad and perpetual license over inputs and outputs, and even your name and profile picture. Vague prohibitions on whatever may harm Z.ai’s "interests" or even the "national interests" of any country. Vague prohibitions on "disturbing" or "inappropriate" content, whatever that is. Vague prohibitions on discussing Z.ai, even my posting this comment violates it. Can ban you…

Are their TOS significantly more vague or restrictive than OpenAI or Anthropic’s?

In any case what matters is what is enforced in practice. It will be a mild inconvenience to switch providers on Openrouter.

If Anthropic or OpenAI decide to apply those same arbitrary terms, you are SOL.

Post reply on HN