Live data from Hacker News

Cerebras CS-4

cerebras.ai

111–120 of 281 posts

Re: Cerebras CS-4

#111

Earlier quoted context omitted.

I mean we kinda know the frontier models are multi trillion parameter models. The only open weights that are close to the frontier are that size too

save qwen3.8 27B which is outclassing much larger models and is in spitting distance of the top 10 in https://artificialanalysis.ai/models#intelligence

I wonder why they removed DeepSWE from their incorporates evaluations

Re: Cerebras CS-4

#112
post #95

Earlier quoted context omitted.

not only that, but I was so happy with their GLM 4.8 that they got rid of yesterday :(

What do they do with old ones? Their hardware physically can't run other models right?

The Cerebras hardware is not locked to specific models / model families. Taalas is the company that's etching models into their silicon, locking it to that model forever.

Re: Cerebras CS-4

#113
post #70

Earlier quoted context omitted.

I read it as it is impressive because smaller models 2.5T are squeezing similar returns as 10T models despite being 1/4th size not that there beyond 2T today the number or parameters do not have much meaning

or the latest qwen3.8 27B doing so well at ~1/100 the size of K3

What about general knowledge you can get out of it before hallucinations start?

Re: Cerebras CS-4

#114
post #107

OpenAI needs to immediately move to acquire Cerebras. Nvidia's extreme margin is the opportunity for OpenAI's cost reduction. Buying Cerebras would pay for itself and they should take all of its future production (after filling required contracts). Right now China's models have no silicon moat. Cerebras as a drastic speed-up / cost-reduction potential, can assist in building a competitive moat. And every time a Cereb…

Do you know about Jalapeno?

^

OpenAI is partnering with Cerebras while simultaneously investing in their own silicon play. Hedged bets.

After sitting thru their keynote today, it makes sense. The main throughput speedups they tout are an obvious evolution of the GPU that all companies will be building in the next year. Wafer-scale interconnected memory and compute is just going to beat out mountains of network cabling any day on both cost and performance metrics.

Re: Cerebras CS-4

#115
post #113

Earlier quoted context omitted.

or the latest qwen3.8 27B doing so well at ~1/100 the size of K3

What about general knowledge you can get out of it before hallucinations start?

It did OK on schlongbench v1.0 (test of a specific niche word that doesn't make it into smaller LLMs) but it sure does love to count words

https://pastes.io/r8F1AY8h

Re: Cerebras CS-4

#116
post #52

It would be even better if a version available to individual users were released soon.

needs an sla that says power will never ever ever go out or else you will have a useless shattered plate of silicon.

Huh, why would it shatter if the power goes out?

Re: Cerebras CS-4

#117
post #113

Earlier quoted context omitted.

or the latest qwen3.8 27B doing so well at ~1/100 the size of K3

What about general knowledge you can get out of it before hallucinations start?

I do not rely on any LLM of any size for general knowledge baked into the weights, they all hallucinate and that is the wrong way to hold them imo

I think there is some merit in that smaller models cannot memorize so much of the training data, i.e. that they are less likely to do copyright infringement, and by analogy not having memorized SDK / API surfaces that have since changed from the training data

Re: Cerebras CS-4

#118
post #111

Earlier quoted context omitted.

save qwen3.8 27B which is outclassing much larger models and is in spitting distance of the top 10 in https://artificialanalysis.ai/models#intelligence

I wonder why they removed DeepSWE from their incorporates evaluations

They didn't afaict https://artificialanalysis.ai/agents/coding-agents?coding-ag...

It seems it takes some time to run a new model on all the benchies, not sure they run all models on all of them either

Re: Cerebras CS-4

#119

Earlier quoted context omitted.

Is CUDA still a moat? Are we not at the point where frontier models can reimplement software stacks, given you throw enough tokens at the problem.

You just proved that AI cannot currently do that

It can definitely create a software stack for you if you hold it right, but the software stack supported by a trillion dollar company with decades of expertise, that also uses AI to improve its stack is probably gonna be better.

Re: Cerebras CS-4

#120

Earlier quoted context omitted.

Hence why taalas was one of the best strategic acquisitions of the year. I'm honestly baffled they were not acquired by somebody else (sorry AMD).

Taalas will be one of the great disaster investments of the early AI era. It'll be a near total write-down. The absolute worst market time to etch a model to a chip is right now (very rapid iteration). There is no scenario where they can keep up. The Taalas approach will be viewed as comically foolish within just a few years. Cerebras will win in terms of approach. It's 1998: hey, I can drastically speed up your web…

I just want to but hardware so I can run a model at home that is fast. I don't see myself installing a server that burns almost two hundred kilowatts but maybe a card which runs a 27B Qwen...
Post reply on HN