Live data from Hacker News

Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

bloomberg.com

101–110 of 148 posts

Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

#102
post #6

I'd be interested to know what was going on with it during the public test as there were numerous reports of it improving considerably at tasks it was asked to do early on in the test compared to later in it.

Two potentials from my pov: 1. Just variance in pass@K. If you prompt any model multiple times you'll see a large variance. N=1, but I find chinese open source models have a higher variance than higher-RL'd models like fable/opus. 2. They legitimately shipped a new RL checkpoint over the 7 days, which I find hard to believe. I am leaning towards 1.

Or 3, they find some bug/regression in their pipeline; maybe they didn't quant parts of a model properly, maybe their inference engine had a bug, maybe some pinned MoE expert wasn't pinned, etc...

That's very plausible to have, identify, and fix in a day; especially when you get community feedback in the wild.

Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

#103
post #90
post #60

Earlier quoted context omitted.

FWIW, the way GLM-5.2 (and 5.3) talk is clearly claude, so it is for sure also trained using distillation. The metric used there is me screaming at my screen per operating hours. Does it matter? IMO not really. Weights are open after all. (Or.. soon at least for 5.3)

With the amount of Claudish on the internet now, and in source code repositories (how many Claudish README.mds have you seen?), you don't have to make a single API call to end up with a model that talks like Claude. And critically, like contracts in general, Anthropic's terms of service is only binding upon the user/counterparty. So even if a company say specifically sought out 'claude-like' content, and claude code…

[deleted]

Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

#104

Rather than a pelican, for fun I showed it a couple of screenshots from Niu Lai and asked it to create an SVG inspired by the images. I explained a little about how the movie had been made by a mother & son team, initially derided but then went on to surprise cult box office success. It came up with this: https://x.com/syneryder/status/2091978367579156569/photo/1 Created in a single turn - but technically not a "one-…

> I'm enjoying working with Ox in a way that I'm just not enjoying talking to the 5.0 Anthropic models. That’s very valid, but right now every other model I use is easier to talk to than Opus 5.0 Opus 5.0 has an impenetrable way of communicating. I can parse it, but it takes so much more work than it should.

Yeah, that's a fair point. "More intelligible than Adriano Celentano in Prisencolinensinainciusol" is not a high bar.

As another comparison, I went back to MiniMax M3 for a while last night. It was significantly faster than Ox, but I felt M3's replies were harder to parse, not quite getting to the point. But I guess I could curb that with some prompts.

It depends if the Ox Alpha pricing is as cheap as was being rumored. If it's competitive with DeepSeek Flash and significantly undercutting Luna, that feels like it will be significant.

Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

#105

I had Ox Alpha working on coding tasks for a couple days non-stop, via OpenRouter and OpenCode Zen. Crush harness. It was able to complete tasks at a level that I'd put between Sonnet and Opus. It makes few mistakes, but is not that smart. The main issue for me, is that it degraded into a doom loop several times. One of them was running the same bash command about a thousand times. The last model I've used that had t…

Doom loops are as much of a model problem as it is deficiency of the harness. I have not seen any other open source harness that deals with them except the one I started because of this obvious gap.

See my other comment with examples where 0x Alpha is working non-stop on various projects with zero problems.

Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

#106
post #20
post #2

Unfortunately I can't find sources other than this for now but this seems to be legit.

> The company on Wednesday confirmed speculation that the Ox Alpha model is a new iteration of its GLM series and said it will release the weights for it tonight, in response to queries by Bloomberg News. Seems legit. It's really hard to know how good it is. So much hype around it.

I mean, you can try it for free.

Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

#107

Earlier quoted context omitted.

Claims about Ox Alpha performing at Fable level were from the social media hype cycle. Everything new in the LLM space brings a wave of influencers hyping it up as a revolutionary leap forward. Don’t forget to like and subscribe to learn more. It is a capable small model, but it’s not frontier level. The interesting part will be seeing the model size, how it responds to quantization, and how fast it runs on the kind…

This influencers are getting paid, it's not coincidential.

They don't have to be getting paid. The natural bias of media is towards laziness and sensationalism (stolen from Jon Stewart, so maybe the same is true about comments).

Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

#109
post #47

Releasing weights is the right move. Keeps them competitive with DeepSeek on the open side.

I had good experience with GLM 5.3, but... Z.AI is the only provider for GLM 5.3 on OpenRouter. I don't see 5.3 on Hugging Face. Not sure if this new model is "full GLM" or something smaller, or if they will like Moonshot AI publish weights but put restrictive license [1], which will again leave Z.AI as single GLM model provider on OpenRouter. [1] https://huggingface.co/moonshotai/Kimi-K3/blob/main/LICENSE

The release date is supposed to be August 28th 2026

Re: Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights

#110

I had Ox Alpha working on coding tasks for a couple days non-stop, via OpenRouter and OpenCode Zen. Crush harness. It was able to complete tasks at a level that I'd put between Sonnet and Opus. It makes few mistakes, but is not that smart. The main issue for me, is that it degraded into a doom loop several times. One of them was running the same bash command about a thousand times. The last model I've used that had t…

[dead]
Post reply on HN