Earlier quoted context omitted.
The car industry is also a trillion $$ market in the US. I don't see why that would go any differently from the Chinese cars ban.
You wouldn't download a car, would you?
GLM-5.3: Frontier coding with emergent cyber capabilities
201–210 of 626 posts
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#202Apparently they are scanning OSS and popular software at scale and disclosing the vulnerabilities they found: https://cvd.z.ai/ Most of these are under embargo, but it seems there are a lot of CVE here from a wide range of popular software, many considered critical or high. I understand the argument of "people are not actively looking", but isn't the cost for such a scan getting lower by the week, and Anthropic's Pro…
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#203Earlier quoted context omitted.
OpenAI and Anthropic are both seeking trillion IPOs, while Chinese labs are pumping out open-weight models that are free for US providers to host and monetize. These Chinese models cost less of US SOTA models to run, even if they are less capable. Providers can just run them, offer cheap tokens, and pocket the margin. I just don't see how you justify a trillion valuation for US AI labs when the underlying models are…
US investors are desperate for the next hypergrowth opportunity. From what I can tell the US economic strategy is to outgrow its debt.
All investors.
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#204Earlier quoted context omitted.
> Providers can just run them, offer cheap tokens, and pocket the margin. There’s an assumption that you can spin up the infra and acquire customers within that margin
Which is not unreasonable. Just hosting it in the EU and promising not to retain / sell the data let's you charge a healthy extra and compete in many areas other players can't.
It's been a few years. Has anyone done this successfully yet?
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#205Earlier quoted context omitted.
And with that also gain institutional knowledge, skill up your workers and attract talent that wants to work on this stuff. All boils down to short-term/long-term thinking.
This. People WANT to work on this stuff. And having skilled workers is a precious advantage.
For our own model training we needed to do some large scale translation tasks of a large dataset (1M or so documents, 10 or so target languages), running full-size NLLB on-prem saved us an absurd amount of money vs Google Translate API.
(For reference doing 1M target docs into a single language in Google Translate API is roughly $120k list price. You can run full size NLLB on an 48GB NVIDIA A600 and the major difference for us was speed, but for this task time to completion wasn’t an issue.)
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#206Their own hardness (ZCode) seems to be a GUI, which doesn't work for me. They say they support other harnesses. However, it seems like I can inject the plan into other harnesses, like Claude Code[0]. Does anyone who's been using GLM models for a while have a strong feeling for if it does better in some harnesses than others, or should I just use my favourite harness? 0: https://docs.z.ai/devpack/tool/others
$ ANTHROPIC_BASE_URL="https://api.z.ai/api/anthropic" ANTHROPIC_AUTH_TOKEN="zai-api-key" claude --dangerously-skip-permissions
Lately I use Zai in omp (oh-my-pi). It's listed built-in provider can be selected without configs shenanigans. Fits in the overall setup e.g. can select GLM-5.2 (now 5.3), and assign it role [plan] or [advisor]. I got reminded now of glm-5v-turbo. Think that 'v' was for vision. Assigned it role [vision] in omp now, let's see what happens. :-)
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#207Earlier quoted context omitted.
OpenAI and Anthropic are both seeking trillion IPOs, while Chinese labs are pumping out open-weight models that are free for US providers to host and monetize. These Chinese models cost less of US SOTA models to run, even if they are less capable. Providers can just run them, offer cheap tokens, and pocket the margin. I just don't see how you justify a trillion valuation for US AI labs when the underlying models are…
It is impossible to justify the absurd private valuations they have given themselves in collusion with investors. I wish they had tried to IPO because then we’d see the judgement of the market on this. But that’s why they didn’t this year. How long can they keep up the charade that their models are uniquely valuable and on the path to AGI?
What's the collusion?
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#208Earlier quoted context omitted.
Looks like they're going for good PR now, to avoid smearing by the "Western" models. Smart!
I'd love to live in a society where people and corporations do good things for PR.
Maybe so, but I'm not sure I'd like to live in China of all places. (Don't get me wrong. Lotta places I'd like to visit if I ever got the chance, and China's on that list, but to live there? I don't think so.) Maybe one of the Nordic countries?
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#209Earlier quoted context omitted.
Do you actually believe this?
The CIA ran one of the world's largest cryptography companies, for DECADES[1]. Are you truly so naive that you believe intelligence agencies that have more to gain from stifling the discovery of vulnerabilities they know of and use wouldn't do so? [1] https://www.washingtonpost.com/graphics/2020/world/national-...
I'm curious: to any professional vulnerability researchers reading this, what do you think?
Re: GLM-5.3: Frontier coding with emergent cyber capabilities
#210Earlier quoted context omitted.
No time like the present to pull out and reduce your exposure. I brought this up in my employer's forums 4 months ago and honestly it's been clear even before then. In particular, the upcoming IPOs of both oAI and Anthropic will likely be disastrous for the public - the floor is falling from under them and I don't know if they can be scrappy and work with fewer resources - their internal culture may not support this.…
Yes buuuut…. I do quite a bit of day trading (maybe closer to scalping) for the first few hours the market is open, everyday. Anecdotally: despite everyone knowing its valuation was ridiculous, I rode that SpaceX train pretty hard and made a pretty penny. I close-out all my positions by end-of-trading everyday… so when the day came when there was a very clear and very scary indicator during early trading hours, quick…