Cerebras launches Qwen3-235B, achieving 1.5k tokens per second
1–10 of 160 posts
Re: Cerebras launches Qwen3-235B, achieving 1.5k tokens per second
#2Re: Cerebras launches Qwen3-235B, achieving 1.5k tokens per second
#3Re: Cerebras launches Qwen3-235B, achieving 1.5k tokens per second
#4Re: Cerebras launches Qwen3-235B, achieving 1.5k tokens per second
#5Quantization?
Re: Cerebras launches Qwen3-235B, achieving 1.5k tokens per second
#6Very impressive speed. With a context window of 40K however, usability is limited.
Re: Cerebras launches Qwen3-235B, achieving 1.5k tokens per second
#7Very impressive speed. With a context window of 40K however, usability is limited.
> Cerebras Systemstoday [sic] announced the launch of Qwen3-235B with full 131K context support on its inference cloud platform
Then later:
> Cline users can now access Cerebras Qwen models directly within the editor—starting with Qwen3-32B at 64K contexton the free tier. This rollout will expand to include Qwen3-235B with 131K context
Not sure where you get the 40K number from.
Re: Cerebras launches Qwen3-235B, achieving 1.5k tokens per second
#8Re: Cerebras launches Qwen3-235B, achieving 1.5k tokens per second
#9Very impressive speed. With a context window of 40K however, usability is limited.
The first paragraph contains: > Cerebras Systemstoday [sic] announced the launch of Qwen3-235B with full 131K context support on its inference cloud platform Then later: > Cline users can now access Cerebras Qwen models directly within the editor—starting with Qwen3-32B at 64K contexton the free tier. This rollout will expand to include Qwen3-235B with 131K context Not sure where you get the 40K number from.
Re: Cerebras launches Qwen3-235B, achieving 1.5k tokens per second
#10Very impressive speed. With a context window of 40K however, usability is limited.
The first paragraph contains: > Cerebras Systemstoday [sic] announced the launch of Qwen3-235B with full 131K context support on its inference cloud platform Then later: > Cline users can now access Cerebras Qwen models directly within the editor—starting with Qwen3-32B at 64K contexton the free tier. This rollout will expand to include Qwen3-235B with 131K context Not sure where you get the 40K number from.