Live data from Hacker News

Cerebras Code

cerebras.ai

171–180 of 185 posts

Re: Cerebras Code

#171
post #163

Earlier quoted context omitted.

The Cerebras.ai plan offers a flat fee of $50 or $200. The API price is not a reason to reject the subscription price.

The flat fee is for a fixed max amount of tokens per day. Not requests, tokens.

[deleted]

Re: Cerebras Code

#172
post #77

Earlier quoted context omitted.

This seems to be rate limited by message not token so the lack of cache may matter less

No it’s by token. The FAQ says this: > Actual number of messages per day depends on token usage per request. Estimates based on average requests of ~8k tokens each for a median user. https://cerebras-inference.help.usepylon.com/articles/346886...

How did you find that? Are you sure it applies to Cerebras Code Pro or Max?

Re: Cerebras Code

#173

Anyone get this working in Cursor? I can connect openrouter just fine, but Cerebras just errors out instantly. Same url/key works via curl, so some sort of Cerebras/Cursor compatibility issue.

Same here. Got this msg on the Celebras discord:

> Yeah I filed a ticket with Cursor

> They have problems with OpenAI customization

Re: Cerebras Code

#174

Earlier quoted context omitted.

That’s almost no CO2 emissions at all. Here is a CO2 emissions leaderboard (need to sort by the correct column): https://celebrityprivatejettracker.com/leaderboard/

The number one has 32k which is equivellent of 64,000 commercial transantlantic flight trips (per person). For reference, 2024 had a record flights summer of 140k.

For a moment I thought it might be the presidential plane, which would explain the emissions, but no, for some reason Trump's personal plane is a whole ass Boring 757

Re: Cerebras Code

#175

This has to be a monstrous money loser. If they can maintain this pricing level, and if Qwen3‑Coder is as good as people say then they will have an enormous hit on their hands. A massive money losing hit, but a hit. Very interesting! PS: Did they reduce the context window, it looks like it.

Why?

For $200plan, it has 40M token cap per day, so assuming the API pricing, the max usage per day is $12/day or 360 per month. (Assuming user max-out usage every day or doesn't hit the 1000message limit first)

relatively standard subscription pricing vs API pricing, i believe they are making money from this and counting on people compare this to Claude Code, which is a much more generous offer.

Re: Cerebras Code

#176
post #163

Earlier quoted context omitted.

The Cerebras.ai plan offers a flat fee of $50 or $200. The API price is not a reason to reject the subscription price.

The flat fee is for a fixed max amount of tokens per day. Not requests, tokens.

[deleted]

Re: Cerebras Code

#177

Earlier quoted context omitted.

Claude now does have a weekly limit so if you are able to hit your weekly (undisclosed, dynamic) limit in 2 days, you're unable to use the services for the next 5 days. That is what Cerebras is referencing with no weekly limits. Claude has session count limits, dynamic limits within each session, and now weekly limits on top of all that.

Please read my full comment. Cerebras is jumping on a marketing faux-pas by Anthropic. I say this for the point you bring up about monthly session limits - no one on the Claude subreddit has yet to report being hit by this despite many going way over that. These are checks to deal w/ abusive accounts.

> no one on the Claude subreddit has yet to report being hit by this despite many going way over that

Because it hasn't gone into effect yet: "From August 28, we’ll introduce new weekly limits that’ll mitigate these problems while impacting as few customers as possible." [0]

[0] https://xcancel.com/AnthropicAI/status/1949898514844307953#m

Re: Cerebras Code

#178

Earlier quoted context omitted.

This model is super quantized and the quality isn't great, but that's necessary because just like everyone else except for Nvidia and AMD They shat the bed. They went for super crazy fast compute and not much memory, assuming that models would plateu at a fee billion parameters. Last year 70b parameters was considered huge, and a good place to standardize around. Today we have 1t parameter models and we know it still…

Cerebras doesn't normally quantize the models. Do you have more information about this?

It's FP8 [0]

[0]: https://xcancel.com/CerebrasSystems/status/19513503371867015...

Re: Cerebras Code

#179
post #70

Tried this out with Cline using my own API key (Cerebras is also available as a provider for Qwen3 Coder via via openrouter here: https://openrouter.ai/qwen/qwen3-coder ) and realized that without caching, this becomes very expensive very quickly. Specifically, after each new tool call, you're sending the entire previous message history as input tokens - which are priced at $2/1M via the API just like output tokens.…

Does caching make as much sense as a cost saving measure on Cerebras hardware as it does on mainstream GPU's? Caching should be preferred if SSD->VRAM is dramatically cheaper than recalculation. If Cerebras is optimized for massively parallel compute with fixed weights, and not a lot of memory bandwidth into or out of the big wafer, it might actually make sense to price per token without a caching discount. Could someone from the company (or otherwise familiar with it) comment on the tradeoff?

Re: Cerebras Code

#180

Earlier quoted context omitted.

The number one has 32k which is equivellent of 64,000 commercial transantlantic flight trips (per person). For reference, 2024 had a record flights summer of 140k.

For a moment I thought it might be the presidential plane, which would explain the emissions, but no, for some reason Trump's personal plane is a whole ass Boring 757

I'm surprised there hasn't been dick swinging pressure for some billionaire (the type who cant remember how many billions but net worth probably begins with a 1 due to Benford's law) to get a dreamliner as their private jet.
Post reply on HN