Live data from Hacker News

GLM-5.3: Frontier coding with emergent cyber capabilities

z.ai

201–210 of 626 posts

Re: GLM-5.3: Frontier coding with emergent cyber capabilities

#201
post #194
post #154

Earlier quoted context omitted.

The car industry is also a trillion $$ market in the US. I don't see why that would go any differently from the Chinese cars ban.

You wouldn't download a car, would you?

Most of the money will come from companies/corporations who will be required to buy safe AI. The public will be just banned from buying which might make it hard (ie: site/payment blocked) but not impossible. It could be good enough for the big whales.

Re: GLM-5.3: Frontier coding with emergent cyber capabilities

#202
post #128

Apparently they are scanning OSS and popular software at scale and disclosing the vulnerabilities they found: https://cvd.z.ai/ Most of these are under embargo, but it seems there are a lot of CVE here from a wide range of popular software, many considered critical or high. I understand the argument of "people are not actively looking", but isn't the cost for such a scan getting lower by the week, and Anthropic's Pro…

amazing! huge clusters in code from the 1980s haha

Re: GLM-5.3: Frontier coding with emergent cyber capabilities

#203
post #122

Earlier quoted context omitted.

OpenAI and Anthropic are both seeking trillion IPOs, while Chinese labs are pumping out open-weight models that are free for US providers to host and monetize. These Chinese models cost less of US SOTA models to run, even if they are less capable. Providers can just run them, offer cheap tokens, and pocket the margin. I just don't see how you justify a trillion valuation for US AI labs when the underlying models are…

US investors are desperate for the next hypergrowth opportunity. From what I can tell the US economic strategy is to outgrow its debt.

> US investors are desperate for the next hypergrowth opportunity

All investors.

Re: GLM-5.3: Frontier coding with emergent cyber capabilities

#204

Earlier quoted context omitted.

> Providers can just run them, offer cheap tokens, and pocket the margin. There’s an assumption that you can spin up the infra and acquire customers within that margin

Which is not unreasonable. Just hosting it in the EU and promising not to retain / sell the data let's you charge a healthy extra and compete in many areas other players can't.

> Just hosting it in the EU and promising not to retain / sell the data let's you charge a healthy extra and compete in many areas other players can't.

It's been a few years. Has anyone done this successfully yet?

Re: GLM-5.3: Frontier coding with emergent cyber capabilities

#205
post #119

Earlier quoted context omitted.

And with that also gain institutional knowledge, skill up your workers and attract talent that wants to work on this stuff. All boils down to short-term/long-term thinking.

This. People WANT to work on this stuff. And having skilled workers is a precious advantage.

Still has to break even on the balance sheet, especially at a bootstrapped startup. We actually made most of the financial windfall in translation API fees oddly enough.

For our own model training we needed to do some large scale translation tasks of a large dataset (1M or so documents, 10 or so target languages), running full-size NLLB on-prem saved us an absurd amount of money vs Google Translate API.

(For reference doing 1M target docs into a single language in Google Translate API is roughly $120k list price. You can run full size NLLB on an 48GB NVIDIA A600 and the major difference for us was speed, but for this task time to completion wasn’t an issue.)

Re: GLM-5.3: Frontier coding with emergent cyber capabilities

#206

Their own hardness (ZCode) seems to be a GUI, which doesn't work for me. They say they support other harnesses. However, it seems like I can inject the plan into other harnesses, like Claude Code[0]. Does anyone who's been using GLM models for a while have a strong feeling for if it does better in some harnesses than others, or should I just use my favourite harness? 0: https://docs.z.ai/devpack/tool/others

I've used GLM-s the longest with Claude Code and their Anthropic supplied endpoint. As per their docs

$ ANTHROPIC_BASE_URL="https://api.z.ai/api/anthropic" ANTHROPIC_AUTH_TOKEN="zai-api-key" claude --dangerously-skip-permissions

Lately I use Zai in omp (oh-my-pi). It's listed built-in provider can be selected without configs shenanigans. Fits in the overall setup e.g. can select GLM-5.2 (now 5.3), and assign it role [plan] or [advisor]. I got reminded now of glm-5v-turbo. Think that 'v' was for vision. Assigned it role [vision] in omp now, let's see what happens. :-)

Re: GLM-5.3: Frontier coding with emergent cyber capabilities

#207

Earlier quoted context omitted.

OpenAI and Anthropic are both seeking trillion IPOs, while Chinese labs are pumping out open-weight models that are free for US providers to host and monetize. These Chinese models cost less of US SOTA models to run, even if they are less capable. Providers can just run them, offer cheap tokens, and pocket the margin. I just don't see how you justify a trillion valuation for US AI labs when the underlying models are…

It is impossible to justify the absurd private valuations they have given themselves in collusion with investors. I wish they had tried to IPO because then we’d see the judgement of the market on this. But that’s why they didn’t this year. How long can they keep up the charade that their models are uniquely valuable and on the path to AGI?

> private valuations they have given themselves in collusion with investors.

What's the collusion?

Re: GLM-5.3: Frontier coding with emergent cyber capabilities

#208
post #183
post #175

Earlier quoted context omitted.

Looks like they're going for good PR now, to avoid smearing by the "Western" models. Smart!

I'd love to live in a society where people and corporations do good things for PR.

> I'd love to live in a society where people and corporations do good things for PR.

Maybe so, but I'm not sure I'd like to live in China of all places. (Don't get me wrong. Lotta places I'd like to visit if I ever got the chance, and China's on that list, but to live there? I don't think so.) Maybe one of the Nordic countries?

Re: GLM-5.3: Frontier coding with emergent cyber capabilities

#209
post #185

Earlier quoted context omitted.

Do you actually believe this?

The CIA ran one of the world's largest cryptography companies, for DECADES[1]. Are you truly so naive that you believe intelligence agencies that have more to gain from stifling the discovery of vulnerabilities they know of and use wouldn't do so? [1] https://www.washingtonpost.com/graphics/2020/world/national-...

I believe it is unlikely. (Not because I do not believe NSA is hoarding 0-days, but for many other reasons.)

I'm curious: to any professional vulnerability researchers reading this, what do you think?

Re: GLM-5.3: Frontier coding with emergent cyber capabilities

#210
post #170

Earlier quoted context omitted.

No time like the present to pull out and reduce your exposure. I brought this up in my employer's forums 4 months ago and honestly it's been clear even before then. In particular, the upcoming IPOs of both oAI and Anthropic will likely be disastrous for the public - the floor is falling from under them and I don't know if they can be scrappy and work with fewer resources - their internal culture may not support this.…

Yes buuuut…. I do quite a bit of day trading (maybe closer to scalping) for the first few hours the market is open, everyday. Anecdotally: despite everyone knowing its valuation was ridiculous, I rode that SpaceX train pretty hard and made a pretty penny. I close-out all my positions by end-of-trading everyday… so when the day came when there was a very clear and very scary indicator during early trading hours, quick…

Yes experienced investors will profit from it and leave the general public holding the bag, that's the plan I'm afraid.
Post reply on HN