Live data from Hacker News

Bringing Up DeepSeek-V4-Flash on AMD MI300X

fergusfinn.com

11–20 of 29 posts

Re: Bringing Up DeepSeek-V4-Flash on AMD MI300X

#12
Checked out this company about a year ago and they only offered small models. Now I see they have GLM-fp8/Kimi and DeepSeek V4 Pro. Since workloads are predominantly cached input, I'm surprised to see no separate price for cached input vs uncached. I hope the prices will drop significantly; with these prices you'll end up with thousands in monthly costs quickly. Hopefully more hardware companies will be on the market in the coming years. If the Chinese eventually start competing with the current memory makers, maybe that will help.

Re: Bringing Up DeepSeek-V4-Flash on AMD MI300X

#14
post #12

Checked out this company about a year ago and they only offered small models. Now I see they have GLM-fp8/Kimi and DeepSeek V4 Pro. Since workloads are predominantly cached input, I'm surprised to see no separate price for cached input vs uncached. I hope the prices will drop significantly; with these prices you'll end up with thousands in monthly costs quickly. Hopefully more hardware companies will be on the market…

[flagged]

Re: Bringing Up DeepSeek-V4-Flash on AMD MI300X

#18
post #8

Nice work and thanks for being a customer. (CEO Hot Aisle)

I wish you guys could partner with Modular to get Mojo inference working on your hardware, e.g. https://www.modular.com/models/deepseek-v4-pro

Not sure I understand. If they support MI300x, their self-hosted will run on our hardware.

Re: Bringing Up DeepSeek-V4-Flash on AMD MI300X

#19

Earlier quoted context omitted.

Interesting that you ask that as AMD hits another ATH.

Then you are definitely long on AMD.

More accurately... I'm long on a viable alternative to the current monopoly. We have two OS's for phones (android and ios), there is no reason why we shouldn't have the same for all AI hardware and software. The only one even close, is AMD.
Post reply on HN