Live data from Hacker News

Cerebras Code

cerebras.ai

151–160 of 185 posts

Re: Cerebras Code

#151

> running at speeds of up to 2,000 tokens per second, with a 131k-token context window, no proprietary IDE lock-in, and no weekly limits! I was excited, then I read this: > Send up to 1,000 messages per day—enough for 3–4 hours of uninterrupted vibe coding. I don't mind paying for services I use. But it's hard to take this seriously when the first paragraph claim is contradicting the fine prints.

We're just doing usage-based pricing for our ai devtools product because it's the only way to square the circle of "as much access to an expensive thing as you want, at a reasonable price".

It's harder to set up, lends itself to lower margins, and consumers generally do prefer more predictable/simpler pricing, but so many ai devtools products have pissed their users off by throttling their "unlimited"/plan-based pricing that I think it's now seen as a yellow flag

Re: Cerebras Code

#152
post #34

Earlier quoted context omitted.

1,000 messages per day should be plenty as a daily development driver. I use claude code sonnet 4 exclusively and I do not send more than 1,000 messages per day. However, that is my current understanding. I am certainly not pressing enter 1,000 times! Maybe there are more messages being sent under the hood that I do not realize?

Still not sure if it's 1000 messages or calls though, if messages that's good.

Logically it can only be API call based, as you bring your own IDE plugin. So there's no possibility it's based upon any UI level concept such as top level messages. The subscription wouldn't even necessarily know.

Re: Cerebras Code

#155

Earlier quoted context omitted.

1,000 messages per day should be plenty as a daily development driver. I use claude code sonnet 4 exclusively and I do not send more than 1,000 messages per day. However, that is my current understanding. I am certainly not pressing enter 1,000 times! Maybe there are more messages being sent under the hood that I do not realize?

Your “one enter” press might generate dozens or even hundreds of messages in an agent. Every file read, re-read, read a bit more, edit, whoops re-edit, ls, grep, etc etc counts as a message.

[deleted]

Re: Cerebras Code

#156

Some users who signed up for pro ($50 p.m.) are reporting further limitations than those advertised. >While they advertise a 1,000-request limit, the actual daily constraint is a 7.5 million-token limit. [1] Assumes an average of 7.5k/request whereas in their marketing videos they show API requests ballooning by ~24k per request. Still lower than the API price. [1] https://old.reddit.com/r/LocalLLaMA/comments/1mfeazc…

shocking..

Re: Cerebras Code

#157

Earlier quoted context omitted.

Claude Code does not have a daily limit, it has 5 hour windows that reset. On the $100 plan it's pretty hard to hit a window limit w/ Sonnet unless you're using multiple/subagents. The $200 is better suited to those who do that or want to use a significant amount of Opus. Also the weekly limit selling point is silly - it almost certainly only impacts those who are abusing, ie. running 24/7.

Claude now does have a weekly limit so if you are able to hit your weekly (undisclosed, dynamic) limit in 2 days, you're unable to use the services for the next 5 days. That is what Cerebras is referencing with no weekly limits. Claude has session count limits, dynamic limits within each session, and now weekly limits on top of all that.

Please read my full comment.

Cerebras is jumping on a marketing faux-pas by Anthropic. I say this for the point you bring up about monthly session limits - no one on the Claude subreddit has yet to report being hit by this despite many going way over that. These are checks to deal w/ abusive accounts.

Re: Cerebras Code

#158
post #49

Their hardware is incredible. Why aren’t more investors lining up for this in this environment?

Contradictions do not exist. Whenever you think that you are facing a contradiction, check your premises. You will find that one of them is wrong.

Kurt Gödel would like a word

Re: Cerebras Code

#159
post #49

Their hardware is incredible. Why aren’t more investors lining up for this in this environment?

This model is super quantized and the quality isn't great, but that's necessary because just like everyone else except for Nvidia and AMD They shat the bed. They went for super crazy fast compute and not much memory, assuming that models would plateu at a fee billion parameters. Last year 70b parameters was considered huge, and a good place to standardize around. Today we have 1t parameter models and we know it still…

Cerebras doesn't normally quantize the models. Do you have more information about this?

Re: Cerebras Code

#160
post #145
post #137

Earlier quoted context omitted.

At this point I'm afraid to ask, but I will do I anyways: How do Claude's rate limits actually work? I'm not a Pro/Max5/Max20 subscriber, only light API usage for Anthropic - so it's likely that I don't really understand the limits there. For example, community reports that Anthropic's message limit for Max 5 translates to roughly 88k token per 5-hour window (there's variance, but it's somewhere in this 80-120k ballp…

With Anthropic's Claude subscriptions - while many people appear to use tokens as an idea of the usage limit, I doubt that's what is really used by Anthropic. Why do I say this? Well, there are multiple models, Haiku, Sonnet and Opus, we all know that Opus is the most expensive and burns through the usage limit of the subscription the fastest of all. I'd theorise that Anthropic have some kind of internal credit value…

It's probably best to look at it as credit based, which map to a certain scale of particular tokens (ie. an Opus token takes 5x the credits of a Sonnet token).
Post reply on HN