GLM-5.3 is now open-weight
161–170 of 296 posts
Re: GLM-5.3 is now open-weight
#162Earlier quoted context omitted.
I have just built an Epyc with 512gb DDR4 3200 RAM for a "reasonable" price and I'm hoping to have a setup with GLM as the architect and Qwen 27b/Next Flash as the implementer. This is 1/5 of the price of the Mac, but also probably 1/5 of the speed lol.
Honestly I suspect neither of them will be performing terribly well but with DDR4 3200 RAM I wonder if you'll be counting tokens per second or seconds per token. I mean, you do at least get a lot of memory channels at least, compared to consumer PCs. I am curious to hear what performance you get, I feel there is not enough information out there on what different setups manage to eek out.
Re: GLM-5.3 is now open-weight
#163Earlier quoted context omitted.
I have just built an Epyc with 512gb DDR4 3200 RAM for a "reasonable" price and I'm hoping to have a setup with GLM as the architect and Qwen 27b/Next Flash as the implementer. This is 1/5 of the price of the Mac, but also probably 1/5 of the speed lol.
It’s not unified ram? I.e VRAM so it will struggle
Re: GLM-5.3 is now open-weight
#164Earlier quoted context omitted.
I have just built an Epyc with 512gb DDR4 3200 RAM for a "reasonable" price and I'm hoping to have a setup with GLM as the architect and Qwen 27b/Next Flash as the implementer. This is 1/5 of the price of the Mac, but also probably 1/5 of the speed lol.
Depending on which Epyc you got it might be slower than 1/5 of the speed.
Re: GLM-5.3 is now open-weight
#165Re: GLM-5.3 is now open-weight
#166Re: GLM-5.3 is now open-weight
#167h/t to DeepInfra for being the first 3rd party provider for it on OpenRouter ( https://openrouter.ai/z-ai/glm-5.3?endpoint=b711bea7-3994-49... ).
Re: GLM-5.3 is now open-weight
#168Is it possible to fine tune this model and unlock / extend its cybersecurity capabilities? I'm scared that maybe we are not ready for an open-weight model with high cybersecurity skills.
The insufferable gatekeeping of the US companies is actively contributing to computer insecurity at this point.
Re: GLM-5.3 is now open-weight
#169Does this mean it'll be on Bedrock soon? I hear great things about this model but I want AWS data handling practices...
Re: GLM-5.3 is now open-weight
#170Earlier quoted context omitted.
I think it'd get you less than a $20 sub to any of the big three. I've used it on OpenRouter and found it kind of expensive for the results, but that might change now that it's open weight and other providers can host it/compete with Z.ai. For the work I did with it, I would've rather used DeepSeek V4 Flash just because it's more economical and still gives good results IMO. Z.ai does have their own subscription, but…
> but I haven't used it because their privacy policy was pretty buns last time I checked. What did you find objectionable? I looked at it when I subscribed almost a year ago and I was fine with it (e.g. they don't train on your API inputs).