Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
101–110 of 153 posts
Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
#102I'd be interested to know what was going on with it during the public test as there were numerous reports of it improving considerably at tasks it was asked to do early on in the test compared to later in it.
Two potentials from my pov: 1. Just variance in pass@K. If you prompt any model multiple times you'll see a large variance. N=1, but I find chinese open source models have a higher variance than higher-RL'd models like fable/opus. 2. They legitimately shipped a new RL checkpoint over the 7 days, which I find hard to believe. I am leaning towards 1.
That's very plausible to have, identify, and fix in a day; especially when you get community feedback in the wild.
Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
#103Earlier quoted context omitted.
FWIW, the way GLM-5.2 (and 5.3) talk is clearly claude, so it is for sure also trained using distillation. The metric used there is me screaming at my screen per operating hours. Does it matter? IMO not really. Weights are open after all. (Or.. soon at least for 5.3)
With the amount of Claudish on the internet now, and in source code repositories (how many Claudish README.mds have you seen?), you don't have to make a single API call to end up with a model that talks like Claude. And critically, like contracts in general, Anthropic's terms of service is only binding upon the user/counterparty. So even if a company say specifically sought out 'claude-like' content, and claude code…
Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
#104Rather than a pelican, for fun I showed it a couple of screenshots from Niu Lai and asked it to create an SVG inspired by the images. I explained a little about how the movie had been made by a mother & son team, initially derided but then went on to surprise cult box office success. It came up with this: https://x.com/syneryder/status/2091978367579156569/photo/1 Created in a single turn - but technically not a "one-…
> I'm enjoying working with Ox in a way that I'm just not enjoying talking to the 5.0 Anthropic models. That’s very valid, but right now every other model I use is easier to talk to than Opus 5.0 Opus 5.0 has an impenetrable way of communicating. I can parse it, but it takes so much more work than it should.
As another comparison, I went back to MiniMax M3 for a while last night. It was significantly faster than Ox, but I felt M3's replies were harder to parse, not quite getting to the point. But I guess I could curb that with some prompts.
It depends if the Ox Alpha pricing is as cheap as was being rumored. If it's competitive with DeepSeek Flash and significantly undercutting Luna, that feels like it will be significant.
Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
#105I had Ox Alpha working on coding tasks for a couple days non-stop, via OpenRouter and OpenCode Zen. Crush harness. It was able to complete tasks at a level that I'd put between Sonnet and Opus. It makes few mistakes, but is not that smart. The main issue for me, is that it degraded into a doom loop several times. One of them was running the same bash command about a thousand times. The last model I've used that had t…
See my other comment with examples where 0x Alpha is working non-stop on various projects with zero problems.
Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
#106Unfortunately I can't find sources other than this for now but this seems to be legit.
> The company on Wednesday confirmed speculation that the Ox Alpha model is a new iteration of its GLM series and said it will release the weights for it tonight, in response to queries by Bloomberg News. Seems legit. It's really hard to know how good it is. So much hype around it.
Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
#107Earlier quoted context omitted.
Claims about Ox Alpha performing at Fable level were from the social media hype cycle. Everything new in the LLM space brings a wave of influencers hyping it up as a revolutionary leap forward. Don’t forget to like and subscribe to learn more. It is a capable small model, but it’s not frontier level. The interesting part will be seeing the model size, how it responds to quantization, and how fast it runs on the kind…
This influencers are getting paid, it's not coincidential.
Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
#108Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
#109Releasing weights is the right move. Keeps them competitive with DeepSeek on the open side.
I had good experience with GLM 5.3, but... Z.AI is the only provider for GLM 5.3 on OpenRouter. I don't see 5.3 on Hugging Face. Not sure if this new model is "full GLM" or something smaller, or if they will like Moonshot AI publish weights but put restrictive license [1], which will again leave Z.AI as single GLM model provider on OpenRouter. [1] https://huggingface.co/moonshotai/Kimi-K3/blob/main/LICENSE
Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
#110I had Ox Alpha working on coding tasks for a couple days non-stop, via OpenRouter and OpenCode Zen. Crush harness. It was able to complete tasks at a level that I'd put between Sonnet and Opus. It makes few mistakes, but is not that smart. The main issue for me, is that it degraded into a doom loop several times. One of them was running the same bash command about a thousand times. The last model I've used that had t…