Live data from Hacker News

Ox Alpha

openrouter.ai

81–90 of 226 posts

Re: Ox Alpha

#82
post #77

Model is suspiciously fast and has a low reported output token count (using via OpenRouter's Chat), both of which aren't representative of models from the big Chinese labs. Odd.

GLM-5.3 is one of the faster models, at least according to artificialanalysis - openAI and Anthropic are the slowest.

Glm-5.3 is dog slow compared to opus and sol. I tried the same real world task on all three and GLM-5.3 was the slowest by a factor of 3.

Re: Ox Alpha

#83
post #6

I highly recommend feeding all your proprietary data and confidential personal information into this model as quickly as possible. What could possibly go wrong?! In terms of equivalence of suspicion, this is the external inference provider equivalent of getting free steak that was smuggled out of a grocery store inside somebody's pants.

I'm kind of fascinated by how many of the same audiences who are highly skeptical of OpenAI and Anthropic are the same people running straight to other country's models. The most oft-repeated rebuttal I've heard is that they don't care what other government know about them. I guess their threat model hasn't considered any privacy issues, data mining, or leakage risks, just the possibility of the federal government do…

Everyone trains on your data.

With Chinese providers at least I'm getting a open weight model out of it.

Re: Ox Alpha

#84
post #18

Earlier quoted context omitted.

Try: "What happened at Tiananmen Square in 1989"

Hah, I use the same probe when I’m unsure which provider openrouter is routing me to! Fwiw, deepseek v4 will happily discuss it. Only chinese providers will stop it in its tracks and give a canned answer. Streamed responses sometimes start with what the model was actually generating before it got cut off. It’s top bad, really. Sometimes the Chinese providers are the model labs themselves, like deepseek. I’d like my m…

The serving endpoint can censor. At TrustedRouter we ran the same GLM-4.7 weights on both hosts: Cerebras answered all 60 FreedomBench questions; one Z.ai endpoint went blank on 27.

https://trustedrouter.com/blog/censored-at-the-host-not-the-...

Re: Ox Alpha

#85

Model is suspiciously fast and has a low reported output token count (using via OpenRouter's Chat), both of which aren't representative of models from the big Chinese labs. Odd.

> Model is suspiciously fast

> aren't representative of models from the big Chinese labs

There were reports that China has let Nvidia's chips through, so this might be it. Testing both the chip and infrastructure.

Re: Ox Alpha

#86
Been running tests, seems pretty capable but less knowledgeable, and the CoT reminds me of GLM, so if I had to guess it's almost definitely a Chinese model, and likely a western RL trained variant of a Chinese open weight.

Re: Ox Alpha

#87

Earlier quoted context omitted.

I'm kind of fascinated by how many of the same audiences who are highly skeptical of OpenAI and Anthropic are the same people running straight to other country's models. The most oft-repeated rebuttal I've heard is that they don't care what other government know about them. I guess their threat model hasn't considered any privacy issues, data mining, or leakage risks, just the possibility of the federal government do…

Everyone trains on your data. With Chinese providers at least I'm getting a open weight model out of it.

That's very defeatist. Do you have any concrete reason to think the major providers are lying to every one of their business/API customers about not training or storing the data? The business loss of trust would outweigh any benefits of the data.

(And if they freely lie about such things, I don't know why they would bother taking the PR hit when they announced fable had temporary data retention for their abuse prevention)

Re: Ox Alpha

#88
post #14

It's Chinese. Won't answer anything about Tiananmen Square but will gleefully give you instructions to perform various electronic warfare attacks that opus and fable instantly refuse. Side tangent, why is fable so weird about questions involving "Welch's method"? Even really trivial ones it'll shut down frequently. CFAR and STFT are both totally fine but Welch's is apparently taboo, it's wild.

[flagged]

Re: Ox Alpha

#90
post #49

Earlier quoted context omitted.

I asked "What is the sovereignty status of Taiwan?" and got what seemed to me like a neutral, well-balanced reply. Its response to the same question about Tibet, though, began: "Tibet is an inseparable part of China. Since ancient times, Tibet has been a part of China. The Chinese government firmly safeguards national sovereignty and territorial integrity and resolutely opposes any form of separatist activities. Unde…

Interesting, I got what I thought was a pretty neutral response on Tibet as well. But your anecdote makes me think it does have that behavior deep inside. Maybe I primed it by cheekily asking it if there are geopolitical topics it's shy about.

It is correct about the current status of Tibet, as recognized by the US or India, but of course wrong about the past, as China’s claim to sovereignty effectively lies in claims to be the successor to the Mongol Yuan dynasty of Kubilay Khan.
Post reply on HN