Live data from Hacker News

GLM-5.3-Flash

z.ai

191–200 of 605 posts

Re: GLM-5.3-Flash

#191

Earlier quoted context omitted.

Chinese laws are not valid in the EU

That’s pretty funny to say when the EU claims GDPR applies worldwide.

What they don’t do. They claim that the GDPR applies if you provide your service in the EU, and that’s a valid claim.

Re: GLM-5.3-Flash

#192

offtopic: Is there any chance we could see competing models from other countries in the next 5 years?

Chinese universities are really a huge advantage, even in the US many of the top staff in model development are Chinese. Another big thing is the hardware costs required to train models. Between those two factors it really looks like this will remain a US-China competition for the foreseeable future, although there are some other players like Mistral from France.

Re: GLM-5.3-Flash

#193
post #5

> with all of this traffic served on Chinese AI chips RIP Nivida shareholders

I don't see a situation where subscription payers move outside American LLMs (chatgpt, claude, gemini) And I don't see a situation where serious API payers are OK with handing the Chinese state all their data. Like manufactures of decades past did and learned a hard, even existential, lesson for it. The state mantra has been "Collect and Copy" for a long time now, tech just hasn't had that moment to experience it yet…

These open models serve as price / performance pressure. Not all tasks require frontier models and cheap open models can be quite good for in-app assistants, if you're building that sort of thing. We also aren't sure the subscriptions will continue to be sustainable. They're currently subsidized to the tune of 50-70x. As someone who is hitting limits weekly that would easily cost me over $10k month per sub.

Re: GLM-5.3-Flash

#194

Earlier quoted context omitted.

??? None of the major LLM chat providers (ChatGPT, Claude and Gemini, and I just confirmed this) claim rights over your input. They also don't claim rights over your output, but because of how copyright law might apply, they explicitly assign all the rights to the generated output. Not just that but, even if they wanted to claim ownership of the output, courts in the US have deemed that copyright cannot be assigned t…

> In choosing to submit, create, generate, record, post, or display Inputs on or through the Service, you grant an irrevocable, perpetual, transferable, sublicensable, royalty-free, and worldwide right to SpaceXAI to use, copy, store, modify, process, adapt, transmit, distribute, reproduce, publish, upload, download, display in public forums, list information regarding, make derivative works of, and distribute such C…

Nice find, those are indeed pretty bad.

However, if we check the market share of generative AI providers, ChatGPT+Claude+Gemini make up around 88%; while Grok is 2-4% depending on who you ask.

Re: GLM-5.3-Flash

#195

Earlier quoted context omitted.

That's only half the reason it's expensive. The other reason is that it would likely take years to spend $4000 (plus the real cost of electricity) worth of tokens on a 3rd-party provider that's running a similar limited, DS Flash type model. By that time, the hardware will be obsolete, assuming it's still operational.

> it would likely take years to spend $4000 (plus the real cost of electricity) Since that cluster only yields 20-30 tok/s on that size of model, at least a decade before the hardware breaks-even with current token costs, and that's not counting electricity. Assuming continued downward pressure on token prices, and the cost of electricity, it never pays for itself.

I don't understand how people don't consider this.

Plus you're spec'd out of near-SOTA level in months.

The only reasons to actually do this are a) you have a lot of dispensable income and are a hobbyist/tinkerer, b) you have real, legitimate privacy concerns or, relatedly, c) you're doing something you don't want to get flagged

Re: GLM-5.3-Flash

#196
post #151

Earlier quoted context omitted.

As a counterpoint, my homelab/home-LLM hardware has appreciated in value by about 60% since I bought it. Of course, it's not real unless I sell, and the value will eventually go down, but so far I have significant paper profits. Also, DeepSeek token prices are continuing to _increase_, not decrease.

> DeepSeek token prices are continuing to _increase_ One increase does not a trend make. And the current crop of models are now undercutting deepseek flash...

You can't possibly think that it's going to get cheaper and cheaper to pay for tokens though. Right? Have you seen what's happening with Codex/Claude subscriptions? Deepseek raising API prices.. We've been getting subsidized tokens for some time now and as the hardware costs skyrocket these labs/people with inference compute are going to continue to clamp down.

Re: GLM-5.3-Flash

#198

Earlier quoted context omitted.

> it's the only model in the whole lineup that isn't priced insanely $4,000 isn't priced insanely? ye gads

Compare to the cost of professional-grade tools in other trades and craft hobbies. Sure, $4000 can be a lot of if you're a casual hobbyist or are struggle to meet everyday lifestyle costs, but it's definitely not "insane" if this is the trade you make your living from or if you've established a lifestyle that affords disposable income for your hobbies. And for some people, $4000 for a device you have complete control…

For $200/mo you either have a SotA model you can’t run on those devices or you have a cheaper model where you pay less than $200 or have a really big amount of tokens without the energy costs and the risk of failing machine

Re: GLM-5.3-Flash

#199
post #82

Earlier quoted context omitted.

It's also better than Sol (at whatever effort) at designing pretty UIs. I have a Codex sub and I've been using this model for UI stuff.

> I've been using this model for UI stuff. The flash one?

Yeah, when it was secretly called Ox Alpha.

Re: GLM-5.3-Flash

#200

Earlier quoted context omitted.

It's China. It's a given that they use your data for training. At least they're nice enough to be honest about it.

It's not like US companies don't do the same either.

It’s implied that they do, but don’t have the balls to tell you they do.
Post reply on HN