Live data from Hacker News

GLM-5.3-Flash

z.ai

211–220 of 605 posts

Re: GLM-5.3-Flash

#212
post #204

Earlier quoted context omitted.

> I've been using this model for UI stuff. The flash one?

Yes. I have no UI experience, and wanted a model that could produce something good without me telling it how anything should look like. My prompt was something like: "here's data I have, here's what matters to me, create HTML mockup". All GPT 5.6 models were laughably bad. And I don't want to downplay it - they were just absolutely, objectively horrible. Every single attempt was what I could probably call "if json wa…

Yeah, exact same for me. K3 used to be my go to for UI but it's quite expensive. Ox alpha being so cheap and so comparatively good at design is crazy.

I haven't tried it but I think Qwen Max is also very good at design.

Re: GLM-5.3-Flash

#213

This is going so fast! What a time to be on hackernews: July 16th: The "Kimi K3 moment" - China has caught up to Opus! 4 weeks later: GLM 5.3 - Same performance, but cut the amount of parameters and cost to a third! 12 days later: GLM 5.3 Flash - Almost GLM5.3 performance but cut the parameters in half, cut prices to a fifth and serving on Chinese chips!

And don't forget the coolest part, DeepSeek, Qwen, Z.ai and Moonshot have almost caught up while being open about their research and their model weights. We can mostly speculate about OAI and Anthropic models, nothing else, how fun huh?

Re: GLM-5.3-Flash

#214
post #109

Earlier quoted context omitted.

I get all that. Then alternatives are: - Grok - where I absolutely have 0 trust in X.ai's interst in "pushing humanity forward". - OpenAI and Anthropic - which seem to try to be building the biggest moat they can by pushing to ban open models. And at the same time want to be an Arbiter of what level of intelligence I can use. - Google and Meta - I don't need to talk about the practices of these companies. Yes, the te…

All the American companies you mentioned still follow American law and regulation. Skirting that blatantly has big consequences. Chinese companies do not follow American laws and there are absolutely no consequences for violating it. Moreover, the average American is not even aware of exactly what the legal/judicial environment is like in China. If your code and data is stolen, you can't fly to China and demand justi…

> All the American companies you mentioned still follow American law and regulation. Skirting that blatantly has big consequences. > Chinese companies do not follow American laws and there are absolutely no consequences for violating it.

... lmk when anthropic/openai/spacex/xai are held accountable for anything. Anything at all. Hard to be when you're _writing_ the rules.

Re: GLM-5.3-Flash

#216

Earlier quoted context omitted.

It cuts both ways. A GPU in your basement is a depreciating asset with fixed computing power and consumes electricity. Switching model providers is trivial.

> A GPU in your basement is a depreciating asset All decades prior and up to about a year ago, I would have agreed with you. My Framework Desktop, however has appreciated in value by 75% since I bought it. Will it stay there for a long time? Probably not. But it shows that there are no hard and fast rules about things anymore.

I just bought a Framework Desktop. Would have been nice to get it at the introductory price, or perhaps the new 192gb model refresh they’re now teasing, but I settled and got a 64 gb model. At the time, the 128’s price had already risen again, but the 64’s price was still at a lower price.

64 can still easily do a Qwen 4.8 model, so I’m relatively happy with my purchase… plus, it’s price change has caused it to quickly appreciate in value… so I could sell it if my situation ever turned dire lol

Re: GLM-5.3-Flash

#217
post #5

> with all of this traffic served on Chinese AI chips RIP Nivida shareholders

Ox Alpha is a smaller model and it was running very slowly. Chinese AI accelerators are coming along, but nVidia’s lead is huge.

> and it was running very slowly

... I'm at a loss for words here. It was being served for free. To the entire world.

Re: GLM-5.3-Flash

#218

Earlier quoted context omitted.

I have exactly the same opinion Over the last couple years I’ve had to learn sales and understand the thought process behind this better, and I think I’m beginning to understand it The psychology is that most people aren’t really trying to optimize for productivity (even most people who think they are) on an ROI basis, because their compensation is too decoupled from their actual raw output, and more closely coupled…

These are great points. It's a little off topic but what you bring up is why i advise new grads to spend the first couple years of their career in small eat-what-you-kill companies. I think software devs who start out in large companies get this distorted view that their twice a month direct deposit is just magic and comes from the ether no matter what they do. The whole industry would be better off if everyone start…

Strong agree, but I also think some roles in big companies (for me, infrastructure) or in certain industries (eg trading/finance) can help build the same understanding without as much of the variance/raw exposure to bottom line.

Now that the role of the ticket-cruncher is on the path towards full commoditization, and individuals can move much more quickly (and even more carelessly!), I think product roles will probably shift towards one where developers are more deeply embedded in the product/business process so that they own/understand what to build without as much separation between the decision-making and prioritization of what to build. Or at least, they should.

It was eye opening to me to run the math of "should X people work for Y months on this project to save Z per year?" and realize that in so many cases, the time and effort it would cost to stop "wasting" money on things is WAY more than you could actually save on it. Even "small" projects can very quickly become $1M+ investments in time and resources, and the diminishing returns add up quickly (but also a good way to justify the value of your contributions, when done). But the job only exists if it saves money or makes money...

Re: GLM-5.3-Flash

#219
post #5

> with all of this traffic served on Chinese AI chips RIP Nivida shareholders

Not really. Chinese AI companies were never using NVidia AI chips.

This announcement doesn't really mean anything at all. It means the very few people who are already using Z.ai's API will continue to do so, but the vast majority of money going to Nvidia is through the massive amount of business going to Anthropic, OpenAI, and other western cloud providers and inference providers, who are mostly using NVidia chips for inference.

Also, NVidia chips are still sold out and supply constrained.

Re: GLM-5.3-Flash

#220
post #58
post #5

> with all of this traffic served on Chinese AI chips RIP Nivida shareholders

Not really a brag: it ran like shit. Very slow (~20tps, VERY high latency) and it would timeout all the time. I'm sure the chips are fine, but they clearly didn't have enough capacity for the demand they had (that 100T/day claim was asbolute bs)

[flagged]
Post reply on HN