Live data from Hacker News

Ox-Alpha Is GLM?

dejan.ai

41–50 of 75 posts

Re: Ox-Alpha Is GLM?

#41
post #2

GLM 5.3 and all previous models don't have a vision encoder and can only accept text. Ox-Alpha can accept video and images, so unless Z-ai added a pretty good vision encoder for this model, I don't think so. My money is on Moonshot and this being Kimi K3.5. The measured tps and latency is in-line with K3's tps and latency from Moonshot. MiniMax M3.5 is also possible (but the MiniiMax provider is a lot more performant…

But Moonshot limited signups because they lacked compute.

Re: Ox-Alpha Is GLM?

#42
post #17

If it’s not zhipu then why is it returning errors that zhipu does for other models? Who else would return the exact same errors even if they took a lot of core infra like tokenizer from z?

Ziphu has that many resources to be able to serve capacity for 1 quadrillion tokens per day on Nous portal? My bet is that it's a Composer model from Cursor running on xAI cluster, they already did a Composer based on Kimi-K2.5

Re: Ox-Alpha Is GLM?

#43
> How many words are in the previous message?

Its amazing to me that providers haven't added any sort of masking of the prompt in the thinking traces to avoid prompt extraction via this sort of trivial attack

Re: Ox-Alpha Is GLM?

#44
post #34

Earlier quoted context omitted.

Someone could've trained model on top of GLM. Same way Cognition trained their SWE model on top of Kimi and Cursor did same with their Composer model.

The reasoning levels are the same as GLM 5.3. GLM 5.3 is still not open... I believe it's GLM 5.3 Flash or Air.

Reasoning levels are often just injected system prompts so not a great way to fingerprint models.

Re: Ox-Alpha Is GLM?

#46
post #42
post #17

If it’s not zhipu then why is it returning errors that zhipu does for other models? Who else would return the exact same errors even if they took a lot of core infra like tokenizer from z?

Ziphu has that many resources to be able to serve capacity for 1 quadrillion tokens per day on Nous portal? My bet is that it's a Composer model from Cursor running on xAI cluster, they already did a Composer based on Kimi-K2.5

There's three options here:

- The provider has a massive amount of (unused) hardware. Google or Cursor seem most likely

- The model is extremely efficient, beyond anything we've seen so far

- Whomever made the model has improved the cache efficiency in such a way that it's very cheap to serve. See e.g Deepseeks or Xiaomi caching (pre-price increase)

Re: Ox-Alpha Is GLM?

#47
post #42

Earlier quoted context omitted.

Ziphu has that many resources to be able to serve capacity for 1 quadrillion tokens per day on Nous portal? My bet is that it's a Composer model from Cursor running on xAI cluster, they already did a Composer based on Kimi-K2.5

There's three options here: - The provider has a massive amount of (unused) hardware. Google or Cursor seem most likely - The model is extremely efficient, beyond anything we've seen so far - Whomever made the model has improved the cache efficiency in such a way that it's very cheap to serve. See e.g Deepseeks or Xiaomi caching (pre-price increase)

Option 4: the claimed capacity is not true. Real world usage hasn’t reached anywhere close to it.

Re: Ox-Alpha Is GLM?

#48
post #2

GLM 5.3 and all previous models don't have a vision encoder and can only accept text. Ox-Alpha can accept video and images, so unless Z-ai added a pretty good vision encoder for this model, I don't think so. My money is on Moonshot and this being Kimi K3.5. The measured tps and latency is in-line with K3's tps and latency from Moonshot. MiniMax M3.5 is also possible (but the MiniiMax provider is a lot more performant…

> and all previous models

They have had vision models before just not their flagships

Post reply on HN