Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
11–20 of 119 posts
Re: Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
#12I wish they wouldn't call them "open source models". They aren't open source. They didn't publish the training data. They didn't publish the tools they used to train the model. They published the weights. It's an "open weight model", a term that it seems nearly everyone has agreed is appropriate. Why is this company not using it?
>a term that it seems nearly everyone has agreed is appropriate
Models being considered open source even if the original training code / data is not released also is something almost everyone has agreed to be appropriate.
Re: Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
#13I wish they wouldn't call them "open source models". They aren't open source. They didn't publish the training data. They didn't publish the tools they used to train the model. They published the weights. It's an "open weight model", a term that it seems nearly everyone has agreed is appropriate. Why is this company not using it?
The weights are the preferred form for modifying or integrating with other models. There is no obligation in open source to transitively open source all of the documentation / tools used to create the open source project. >a term that it seems nearly everyone has agreed is appropriate Models being considered open source even if the original training code / data is not released also is something almost everyone has ag…
Open source means open source code. Open weight means a binary file dump, not unlike an exe file. There is nothing open source about it.
Its like having a closed source text editor that censors certain words, and an open source text editor that censors certain words.
The latter can easily be recompiled, the former requires reverse engineering. Both may give you a license to use them freely.
Re: Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
#14I wish they wouldn't call them "open source models". They aren't open source. They didn't publish the training data. They didn't publish the tools they used to train the model. They published the weights. It's an "open weight model", a term that it seems nearly everyone has agreed is appropriate. Why is this company not using it?
The weights are the preferred form for modifying or integrating with other models. There is no obligation in open source to transitively open source all of the documentation / tools used to create the open source project. >a term that it seems nearly everyone has agreed is appropriate Models being considered open source even if the original training code / data is not released also is something almost everyone has ag…
It's the 2nd time I hear this argument and I'm already fed up with it
Is it the preferred way only because training is expensive? It's like saying binaries are the preferred way of modifying program because you can't afford to have a fast enough machine to compile it yourself.
Most people don't have resources to compile their own browser, but what makes some browsers open is access to the source.
Maybe it's the preferred way because the people sharing models are themselves working with weights and no data?
Re: Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
#15Earlier quoted context omitted.
Yeah I think they at least put some light effort into making it readable though. Obviously slop but not quite as bad as most slop articles.
There is barely any effort, a simple GPT5.6 sol pro query rips the post apart.
Re: Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
#16I wish they wouldn't call them "open source models". They aren't open source. They didn't publish the training data. They didn't publish the tools they used to train the model. They published the weights. It's an "open weight model", a term that it seems nearly everyone has agreed is appropriate. Why is this company not using it?
Re: Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
#17Earlier quoted context omitted.
The weights are the preferred form for modifying or integrating with other models. There is no obligation in open source to transitively open source all of the documentation / tools used to create the open source project. >a term that it seems nearly everyone has agreed is appropriate Models being considered open source even if the original training code / data is not released also is something almost everyone has ag…
> There is no obligation in open source to transitively open source all of the documentation / tools used to create the open source project. Open source means open source code. Open weight means a binary file dump, not unlike an exe file. There is nothing open source about it. Its like having a closed source text editor that censors certain words, and an open source text editor that censors certain words. The latter…
I do prefer open weights as being more precise (like, is it even really software that has source code in the first place?) but I feel like at this point the ship has sailed somewhat (though if this is something you're willing to spend your time arguing then... moral support I guess?)
Re: Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
#18Wafer is making themselves synonymous with slop in the inference space. Exaggerated unfair comparisons in all their results, twitter hype posts with alarm emojis etc. > $2.50/GPU-hr for the MI355X, $6.00 for the B300, and $4.25 for the B200. This is not an accurate price comparison for real terms.
Re: Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
#19Wafer is making themselves synonymous with slop in the inference space. Exaggerated unfair comparisons in all their results, twitter hype posts with alarm emojis etc. > $2.50/GPU-hr for the MI355X, $6.00 for the B300, and $4.25 for the B200. This is not an accurate price comparison for real terms.
Re: Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
#20I wish they wouldn't call them "open source models". They aren't open source. They didn't publish the training data. They didn't publish the tools they used to train the model. They published the weights. It's an "open weight model", a term that it seems nearly everyone has agreed is appropriate. Why is this company not using it?