This is absolutely still shy of Sol and Fable, but only just by a hair. Ridiculous results. There's still not a compelling economic reason to drop OpenAI courtesy of the ludicrous reset addiction that's taken place, but it feels like we're on the precipice. How are you all toying with running this kind of thing in a mega quantized way locally? Two weeks out from released weights, but this is still just GLM 5.2 with p…
OpenAI and Anthropic are both seeking trillion IPOs, while Chinese labs are pumping out open-weight models that are free for US providers to host and monetize. These Chinese models cost less of US SOTA models to run, even if they are less capable. Providers can just run them, offer cheap tokens, and pocket the margin. I just don't see how you justify a trillion valuation for US AI labs when the underlying models are…
GLM-5.3: Frontier coding with emergent cyber capabilities
421–430 of 626 posts
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#422I find the performance to size ratio of these models to be way more interesting, selfishly because it makes me bullish on what I'll be able to run on a machine I own over the next few years. The progress is just incredible.
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#423Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#424Apparently they are scanning OSS and popular software at scale and disclosing the vulnerabilities they found: https://cvd.z.ai/ Most of these are under embargo, but it seems there are a lot of CVE here from a wide range of popular software, many considered critical or high. I understand the argument of "people are not actively looking", but isn't the cost for such a scan getting lower by the week, and Anthropic's Pro…
Not with Anthropic's models!
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#425This is absolutely still shy of Sol and Fable, but only just by a hair. Ridiculous results. There's still not a compelling economic reason to drop OpenAI courtesy of the ludicrous reset addiction that's taken place, but it feels like we're on the precipice. How are you all toying with running this kind of thing in a mega quantized way locally? Two weeks out from released weights, but this is still just GLM 5.2 with p…
So exactly the same as Opus 5 and GPT 5.6 Sol. It's all "post-training magic".
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#426I might be just reading my positive bias into that text, but is it possible that it is written less like SV marketing hype trash and more like researchers wrote it? It does feel like it respects both me and my time. Thank you, Z.AI. Amazing what difference it makes when the top of your org are actual university professors.
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#427Earlier quoted context omitted.
You should try a better harness. Try pi, or ohmypi if you want a good OOB experience
I’m in the Claude code harness for everything boat too. What are the alternatives?
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#428Same image->html test as I showed in the Gemini 3.7 flash thread. Note that GLM isn't multimodal, but it still was able to generate something similar-ish by writing a python script to inspect the image and extract elements from it. Original images: https://image.non.io/neonRamenDesigns.webp GLM 5.3 build: https://html.non.io/neonRamenGLM5.3 Opus 5 build for comparison: https://html.non.io/neonRamen For having no visi…
Either way that's pretty impressive.
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#429Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#430Earlier quoted context omitted.
> ... Anthropic's Project Glasswing is supposed to find them quite a while ago? That was my thought too. For all of Anthropic's talk about their "adversaries", it seems Z.AI have been quietly offering fixes for single shot Remote Code Execution flaws in US software (Safari / WebKit) that Apple and Glasswing / Mythos missed, and that Apple would not attribute to GLM.
> That was my thought too. For all of Anthropic's talk about their "adversaries" It’s very likely they found all of them, but that the same happened that happened to Microsoft a couple of decades ago: NSA orders not to disclose / fix them so that they can put it in their collection of unfixed zero days.