Earlier quoted context omitted.
Headline at the top of the Cerebras page linked to by the OP "Cerebras Raises $1.1B Series G at $8.1B Valuation" . If you're going after the AI money gravy train then you need to wave the "we have $n registered users" carrot on your PPT slides for the investors because registered user == monetization opportunity . I'm not defending it. I hate being forced to register for shit when I just want to try it or use the fre…
Well if they give it out for free (aka they pay for it), asking you to register is a reasonable ask. It's not a public service funded by taxpayers.
GPT-OSS 120B Runs at 3000 tokens/sec on Cerebras
11–20 of 31 posts
Re: GPT-OSS 120B Runs at 3000 tokens/sec on Cerebras
#12It’s an absolute beast. I run it via OpenRouter, where I have Groq and Cerebras as the providers. Cheap enough as to be almost free, strong performance, and lightning fast.
Cheap enough for now, but of all the companies selling inference at a loss, Cerebras and Groq are probably losing the most per token. Their hardware is ungodly expensive and its reliance on huge amounts of SRAM bottlenecks how much cheaper it can get, since SRAM density is improving at a snails pace at this point.
Re: GPT-OSS 120B Runs at 3000 tokens/sec on Cerebras
#13Earlier quoted context omitted.
Well if they give it out for free (aka they pay for it), asking you to register is a reasonable ask. It's not a public service funded by taxpayers.
Yes they can ask, but do it at the beginning not the end of the process, this is a dark pattern and fucking annoying.
Re: GPT-OSS 120B Runs at 3000 tokens/sec on Cerebras
#14I absolutely hate it, when a website says "try this" and after you went through the trouble of weiting something comes up with a sign up link first. Makes me leave instantly to never come back.
A week ago I went to a launch party for a product that's supposed to "revolutionize design" (a web app w/ an OAI prompt).
No demo, only like two pictures of the actual product. Founder spent like half an hour giving a speech about the future, etc...
"All of you here will get access to it in a couple weeks."
Couple weeks go by ... I "get access". It's a .dmg, 1) What, I open it, it's not even an app, it's an installer ..., I install it, the app opens up and it's a giant red button that takes you to a website to create an account ...
These guys are completely lost.
Re: GPT-OSS 120B Runs at 3000 tokens/sec on Cerebras
#15Re: GPT-OSS 120B Runs at 3000 tokens/sec on Cerebras
#16I absolutely hate it, when a website says "try this" and after you went through the trouble of weiting something comes up with a sign up link first. Makes me leave instantly to never come back.
This is like declaring that a Ferrari dealership offering you a free test drive in a million dollar art exhibit on wheels is evil for asking for your phone number before handing you the keys. If this was some beat-to-hell, high-mileage used economy car, sure, that would be a pain in the ass, and not worth it. But it's a mistake to place Cerebras into that mental bucket. You don't even need to use real information to…
Re: GPT-OSS 120B Runs at 3000 tokens/sec on Cerebras
#17Earlier quoted context omitted.
Headline at the top of the Cerebras page linked to by the OP "Cerebras Raises $1.1B Series G at $8.1B Valuation" . If you're going after the AI money gravy train then you need to wave the "we have $n registered users" carrot on your PPT slides for the investors because registered user == monetization opportunity . I'm not defending it. I hate being forced to register for shit when I just want to try it or use the fre…
Well if they give it out for free (aka they pay for it), asking you to register is a reasonable ask. It's not a public service funded by taxpayers.
They have other options... rate limiting, serving (more) quantized to non-registered etc. etc.
Re: GPT-OSS 120B Runs at 3000 tokens/sec on Cerebras
#18Earlier quoted context omitted.
Cheap enough for now, but of all the companies selling inference at a loss, Cerebras and Groq are probably losing the most per token. Their hardware is ungodly expensive and its reliance on huge amounts of SRAM bottlenecks how much cheaper it can get, since SRAM density is improving at a snails pace at this point.
Not doubting you but anything to back that up? Happy enough to burn VC money until someone shows up who can run it without losing money, either way.
[1] https://www.sec.gov/Archives/edgar/data/2021728/000162828024...
Re: GPT-OSS 120B Runs at 3000 tokens/sec on Cerebras
#19Earlier quoted context omitted.
Well if they give it out for free (aka they pay for it), asking you to register is a reasonable ask. It's not a public service funded by taxpayers.
Yes they can ask, but do it at the beginning not the end of the process, this is a dark pattern and fucking annoying.
Exactly this.
If you present me with a form and a submit button then I expect the input to go through and a result to be presented.
If you don't want to present me with results before login, then put the form behind the wall too.
Simple.
Re: GPT-OSS 120B Runs at 3000 tokens/sec on Cerebras
#20Earlier quoted context omitted.
This is like declaring that a Ferrari dealership offering you a free test drive in a million dollar art exhibit on wheels is evil for asking for your phone number before handing you the keys. If this was some beat-to-hell, high-mileage used economy car, sure, that would be a pain in the ass, and not worth it. But it's a mistake to place Cerebras into that mental bucket. You don't even need to use real information to…
You didn't get my point at all.