Live data from Hacker News

Qwen 3.8 27B available on Cerebras at 1500 tokens/s

inference-docs.cerebras.ai

21–30 of 238 posts

Re: Qwen 3.8 27B available on Cerebras at 1500 tokens/s

#21
post #10

(Was anyone able to create an account just now? I tried but onboarding falls into a redirect loop) (update: I got my answer. support@ replied and said my email domain is on their blacklist. It was just me (and I've resolved it)).

yeah - used sign in with google

Re: Qwen 3.8 27B available on Cerebras at 1500 tokens/s

#25
I'm saddened that Gemma4 is replaced by Qwen 3.8 on PayGo plan. Gemma4 31B is not coding model but it is excellent at intent understanding and task execution used in agentic software. This just shows that real world dominant usage for llms so far is to code generate. And not to augment business products. They must had barely anyone using Gemma to remove it from that tier.

Re: Qwen 3.8 27B available on Cerebras at 1500 tokens/s

#27
post #2

I used their Coding Plan for a few months. It is genuinely difficult to keep up with the models. The output is so fast. Qwen 3.8 27B is likely one of the strongest models they've hosted so far. Edit: it looks like this is only available on a API token pricing. Does anyone know if they have rolled out prompt caching yet? It used to get pretty expensive for agentic coding tasks with no prompt caching.

Strongest model that they host on the public endpoint. They do a super fast version of GPT 5.6 Sol for OpenAI and have bigger open models on dedicated endpoints.

Re: Qwen 3.8 27B available on Cerebras at 1500 tokens/s

#28
post #2

I used their Coding Plan for a few months. It is genuinely difficult to keep up with the models. The output is so fast. Qwen 3.8 27B is likely one of the strongest models they've hosted so far. Edit: it looks like this is only available on a API token pricing. Does anyone know if they have rolled out prompt caching yet? It used to get pretty expensive for agentic coding tasks with no prompt caching.

The coding plan is gone now right?
Post reply on HN