Live data from Hacker News

GLM-5.3-Flash

z.ai

151–160 of 605 posts

Re: GLM-5.3-Flash

#151

Earlier quoted context omitted.

That's only half the reason it's expensive. The other reason is that it would likely take years to spend $4000 (plus the real cost of electricity) worth of tokens on a 3rd-party provider that's running a similar limited, DS Flash type model. By that time, the hardware will be obsolete, assuming it's still operational.

> it would likely take years to spend $4000 (plus the real cost of electricity) Since that cluster only yields 20-30 tok/s on that size of model, at least a decade before the hardware breaks-even with current token costs, and that's not counting electricity. Assuming continued downward pressure on token prices, and the cost of electricity, it never pays for itself.

As a counterpoint, my homelab/home-LLM hardware has appreciated in value by about 60% since I bought it.

Of course, it's not real unless I sell, and the value will eventually go down, but so far I have significant paper profits.

Also, DeepSeek token prices are continuing to _increase_, not decrease.

Re: GLM-5.3-Flash

#152
post #109

You guys read Z.ai's terms of service, right? Broad and perpetual license over inputs and outputs, and even your name and profile picture. Vague prohibitions on whatever may harm Z.ai’s "interests" or even the "national interests" of any country. Vague prohibitions on "disturbing" or "inappropriate" content, whatever that is. Vague prohibitions on discussing Z.ai, even my posting this comment violates it. Can ban you…

I get all that. Then alternatives are: - Grok - where I absolutely have 0 trust in X.ai's interst in "pushing humanity forward". - OpenAI and Anthropic - which seem to try to be building the biggest moat they can by pushing to ban open models. And at the same time want to be an Arbiter of what level of intelligence I can use. - Google and Meta - I don't need to talk about the practices of these companies. Yes, the te…

All the American companies you mentioned still follow American law and regulation. Skirting that blatantly has big consequences.

Chinese companies do not follow American laws and there are absolutely no consequences for violating it.

Moreover, the average American is not even aware of exactly what the legal/judicial environment is like in China. If your code and data is stolen, you can't fly to China and demand justice in the courts.

Re: GLM-5.3-Flash

#153
The key difference between this and all other GLM models is it's multimodal. You cannot send images to the other GLM models.

Re: GLM-5.3-Flash

#154

Earlier quoted context omitted.

This is not uncommon. Also, it only applies to their chat offering, not the api. OpenRouter also offers the API with ZDR. While shitty, i’d say that its really not that special.

??? None of the major LLM chat providers (ChatGPT, Claude and Gemini, and I just confirmed this) claim rights over your input. They also don't claim rights over your output, but because of how copyright law might apply, they explicitly assign all the rights to the generated output. Not just that but, even if they wanted to claim ownership of the output, courts in the US have deemed that copyright cannot be assigned t…

> In choosing to submit, create, generate, record, post, or display Inputs on or through the Service, you grant an irrevocable, perpetual, transferable, sublicensable, royalty-free, and worldwide right to SpaceXAI to use, copy, store, modify, process, adapt, transmit, distribute, reproduce, publish, upload, download, display in public forums, list information regarding, make derivative works of, and distribute such Content, including anything referenced therein, in any and all media or distribution methods now known or later developed, for any purpose, and to aggregate your User Content and derivative works thereof for any purpose, including but not limited to: (i) maintain and provide the Service; (ii) improve our products and the Service and for our other business purposes, such as data analysis, customer and market research, developing new products or features, or identifying or displaying usage or User Content trends; and (iii) perform such other actions to enforce these Terms, comply with our Privacy Policy, comply with applicable law or governmental, court, and law enforcement requests or requirements or keep our Service safe.

> To the extent the User Content includes a person’s image, likeness, voice, or other similar attributes, you grant SpaceXAI the same rights to use those attributes as part of the User Content as described above. You represent and warrant that you have obtained all rights, licenses, notices, permissions, and consents necessary for SpaceXAI to use that User Content.

https://x.ai/legal/terms-of-service

These are arguably even worse to be honest. Absolute nightmare.

Re: GLM-5.3-Flash

#155

You guys read Z.ai's terms of service, right? Broad and perpetual license over inputs and outputs, and even your name and profile picture. Vague prohibitions on whatever may harm Z.ai’s "interests" or even the "national interests" of any country. Vague prohibitions on "disturbing" or "inappropriate" content, whatever that is. Vague prohibitions on discussing Z.ai, even my posting this comment violates it. Can ban you…

None of that applies if you run it at home. Also 3rd party providers will start serving this pretty soon under different terms.

Re: GLM-5.3-Flash

#156

You guys read Z.ai's terms of service, right? Broad and perpetual license over inputs and outputs, and even your name and profile picture. Vague prohibitions on whatever may harm Z.ai’s "interests" or even the "national interests" of any country. Vague prohibitions on "disturbing" or "inappropriate" content, whatever that is. Vague prohibitions on discussing Z.ai, even my posting this comment violates it. Can ban you…

Give it a couple days, and there will be plenty of other inference companies hosting it. Don't like z.ai's TOS? Use the model on a provider with TOS that you agree with.

Re: GLM-5.3-Flash

#157
post #151

Earlier quoted context omitted.

> it would likely take years to spend $4000 (plus the real cost of electricity) Since that cluster only yields 20-30 tok/s on that size of model, at least a decade before the hardware breaks-even with current token costs, and that's not counting electricity. Assuming continued downward pressure on token prices, and the cost of electricity, it never pays for itself.

As a counterpoint, my homelab/home-LLM hardware has appreciated in value by about 60% since I bought it. Of course, it's not real unless I sell, and the value will eventually go down, but so far I have significant paper profits. Also, DeepSeek token prices are continuing to _increase_, not decrease.

> DeepSeek token prices are continuing to _increase_

One increase does not a trend make. And the current crop of models are now undercutting deepseek flash...

Re: GLM-5.3-Flash

#158
post #7

Weights on HF here: https://huggingface.co/zai-org/GLM-5.3-Flash I decided to take the plunge and get myself four sparks at a decent price (and bought the QSFP cables from AliExpress because they are literally 1/2 the price of Amazon), even knowing Apple was going to release new hardware and there's probably a spark 2 on the horizon. It looks like this is going to be a decent fit for what I need. I've been experiment…

Hopefully you also bought a switch

Re: GLM-5.3-Flash

#159
post #102

Is the actual Z.AI ecosystem good enough to replace the main drivers like Codex and Claude? Because it looks like Z Code is just a Codex fork. Just like the Kimi Code one is. What irks me about this is that the harnesses seem to be just an afterthought here. Don't get me wrong, I love messing around with installing Pi, getting it hooked up with OpenRouter, and just trying all kinds of different stuff, local models, e…

Their list of allowed tools is extensive so just use whatever you want within that list

Think z code gives a token bonus though

Re: GLM-5.3-Flash

#160

The key difference between this and all other GLM models is it's multimodal. You cannot send images to the other GLM models.

I really wish GLM models had vision capabilities. I've worked around that in the past to use a vision MCP in my harness that GLM can call. It is not the same, but it allows the model to query images.
Post reply on HN