Earlier quoted context omitted.
> and it was running very slowly ... I'm at a loss for words here. It was being served for free. To the entire world.
GPT-5.6 Luna is also served for free to the entire world with a tokens per second rate nearly 10X higher. > ... I'm at a loss for words here No need to be so dramatic. I think it's great that they're developing chips, but the whole "RIP nVidia" claim was overly dramatic.
GLM-5.3-Flash
361–370 of 605 posts
Re: GLM-5.3-Flash
#362Chinese labs are so used to manipulating benchmarks to try to flatter inferior models that when they finally have one that's really pretty good I think the official announcement here undersells it. https://deepswe.datacurve.ai/ That's pretty solid. Smarter and cheaper than Luna xhigh, not as smart but less expensive than Luna max. Smashes deepseek v4 flash, and even worse it matches v4 pro at a tiny fraction the cost…
I don't know how anyone can actually use Luna max on ANY real workload. I've had Sol orchestrate a bunch of Luna agents, these agents were explicitly given small chunks of larger objectives and they still filled their entire context windows with just reasoning tokens, until compaction hit, and then reasoning again. I've probably wasted a good 40% of my weekly usage on Luna Max agents just thinking and not writing a s…
Re: GLM-5.3-Flash
#363You guys read Z.ai's terms of service, right? Broad and perpetual license over inputs and outputs, and even your name and profile picture. Vague prohibitions on whatever may harm Z.ai’s "interests" or even the "national interests" of any country. Vague prohibitions on "disturbing" or "inappropriate" content, whatever that is. Vague prohibitions on discussing Z.ai, even my posting this comment violates it. Can ban you…
Yes, the terms are dubious. But they are also reasonably lenient with enforcement. They also don't require persona id verification, witch is wat turned me away from openai.
> They also don't require persona id verification, witch is wat turned me away from openai.
Could be worse. I was dumb enough to verify, only to get rejected for unknown reasons with no retries and no appeals. Had my privacy violated and have nothing to show for it.
Re: GLM-5.3-Flash
#364You guys read Z.ai's terms of service, right? Broad and perpetual license over inputs and outputs, and even your name and profile picture. Vague prohibitions on whatever may harm Z.ai’s "interests" or even the "national interests" of any country. Vague prohibitions on "disturbing" or "inappropriate" content, whatever that is. Vague prohibitions on discussing Z.ai, even my posting this comment violates it. Can ban you…
Are their TOS significantly more vague or restrictive than OpenAI or Anthropic’s? In any case what matters is what is enforced in practice. It will be a mild inconvenience to switch providers on Openrouter. If Anthropic or OpenAI decide to apply those same arbitrary terms, you are SOL.
Yeah, I've compared both. The US companies generally aren't as vague, and they don't claim ownership over inputs and outputs.
Re: GLM-5.3-Flash
#365Re: GLM-5.3-Flash
#366Earlier quoted context omitted.
Not everything is about pure cost. Maybe I don't want to sell my soul supporting the frontier labs because they are straight up pure evil?
I barely see a difference between buying the hardware that feeds (and often colludes with) those labs, at least not as a moral stance. Even if you trained your own model, you'd be committing some of the same sins, paying for the same hardware that drove it, etc. But if you're using some open model, you're standing on the shoulders of the same corrupt giants. I feel like when people say this is due to moral reasons, i…
I also differentiate using their tech and paying them money. I don't think using their models, or perhaps using models derived from them as inherently evil. I just do not want to actually contribute to their bottom line in any way. Even if that means a slower ramp up of AI in general. In my opinion we could move slower.
I understand Nvidia is working with the labs to assist them to buy more hardware through financing and other deals. But ultimately I do not view that as the same as contributing directly to their P&L.
Re: GLM-5.3-Flash
#367Is the actual Z.AI ecosystem good enough to replace the main drivers like Codex and Claude? Because it looks like Z Code is just a Codex fork. Just like the Kimi Code one is. What irks me about this is that the harnesses seem to be just an afterthought here. Don't get me wrong, I love messing around with installing Pi, getting it hooked up with OpenRouter, and just trying all kinds of different stuff, local models, e…
Have been using it as my primary harness for personal work for I'd say 6 months. I recommend everyone create their own harness at least to learn. There are a lot of practical benefits.
Re: GLM-5.3-Flash
#368Is anyone actually tried it in agentic coding (claude code loops)? Are apple silicon macs (M5 Max) capable of working with that model? what was the tps?
Re: GLM-5.3-Flash
#369Re: GLM-5.3-Flash
#370> with all of this traffic served on Chinese AI chips RIP Nivida shareholders
I don't see a situation where subscription payers move outside American LLMs (chatgpt, claude, gemini) And I don't see a situation where serious API payers are OK with handing the Chinese state all their data. Like manufactures of decades past did and learned a hard, even existential, lesson for it. The state mantra has been "Collect and Copy" for a long time now, tech just hasn't had that moment to experience it yet…