Earlier quoted context omitted.
lol find the discord!
yes,sir, any possibility to find 1000pcs or more
I'm super curious what you would use them for.
21–29 of 29 posts
Earlier quoted context omitted.
lol find the discord!
yes,sir, any possibility to find 1000pcs or more
I'm super curious what you would use them for.
Earlier quoted context omitted.
They are all over ebay. https://www.ebay.com/sch/i.html?_nkw=bc-250 I'm super curious what you would use them for.
now many guys want to buy this, I am reseller AMD BC-250,it is popular now
Checked out this company about a year ago and they only offered small models. Now I see they have GLM-fp8/Kimi and DeepSeek V4 Pro. Since workloads are predominantly cached input, I'm surprised to see no separate price for cached input vs uncached. I hope the prices will drop significantly; with these prices you'll end up with thousands in monthly costs quickly. Hopefully more hardware companies will be on the market…
We're kind of known for our low prices - our prices (our main usage is for our high throughput API - the async tier) is significantly below average openrouter prices - but cached prices is coming soon which will lower them even more :)
Nice work! Would DeepSeek V4 Pro on 8xMI300X work with these patches?
Checked out this company about a year ago and they only offered small models. Now I see they have GLM-fp8/Kimi and DeepSeek V4 Pro. Since workloads are predominantly cached input, I'm surprised to see no separate price for cached input vs uncached. I hope the prices will drop significantly; with these prices you'll end up with thousands in monthly costs quickly. Hopefully more hardware companies will be on the market…
Hi! Co-founder of Doubleword here - we've hugely increased the number of models that we offer (partly thanks to work that we've done on hotswapping https://blog.doubleword.ai/fast-sglang-starts . We're kind of known for our low prices - our prices (our main usage is for our high throughput API - the async tier) is significantly below average openrouter prices - but cached prices is coming soon which will lower them e…
Earlier quoted context omitted.
I wish you guys could partner with Modular to get Mojo inference working on your hardware, e.g. https://www.modular.com/models/deepseek-v4-pro
Not sure I understand. If they support MI300x, their self-hosted will run on our hardware.
It's not, which is why it would be nice if they did the actual work (on your hardware).
I would 100% pay $16/hr to run a self-hosted instance, but I won't spend thousands of dollars to (maybe) get it working (my time + the hardware).
Earlier quoted context omitted.
Not sure I understand. If they support MI300x, their self-hosted will run on our hardware.
If it was that easy, I wouldn't have commented. It's not, which is why it would be nice if they did the actual work (on your hardware). I would 100% pay $16/hr to run a self-hosted instance, but I won't spend thousands of dollars to (maybe) get it working (my time + the hardware).
https://docs.modular.com/max/models/
I agree with you though, serving up inference is secret sauce for a lot of teams and not everyone publishes how to do it because of the costs involved in doing so. They need an ROI.