I will get downvoted, but fck it. The ban on these open models is coming within weeks, if not days. As usual, the excuse will be "national security".
Hell, the US doesn't even need to act. I claim the CCP will wise up within 2 years, possibly much much sooner, and ban their own companies from open sourcing to prevent the Americans from acquiring the capabilities. Despite all the nonsense claims of China distilling US models, the reality is that the Americans absolutely do distill these free Chinese models, and distillation when full logprobs are available (i.e. yo…
DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
171–180 of 342 posts
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#172Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#173Somewhat relatedly, how do the economics for Huggingface work? They must be hosting petabytes of models and datasets by now. I have downloaded quite a few “just in case”, only to replace them with the later iteration months later. Does the file hosting actually cost peanuts when you do it yourself and the cloud has shattered my understanding of what it actually costs to deliver so much data?
Bandwidth is really cheap when you run your own infra
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#174[flagged]
Good luck when the last non-Chinese frontier labs will have closed and the CCP will ask to stop sharing models open source.
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#175So GLM 5.2/Gemini 3.6 level intelligence for $0.28/m output. And their updated Pro model coming soon.... Plus a size you can genuinely run at home: Unsloth lossless Q8 at 162GB.
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#176The weights were just released a few minutes ago: https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#177Somewhat relatedly, how do the economics for Huggingface work? They must be hosting petabytes of models and datasets by now. I have downloaded quite a few “just in case”, only to replace them with the later iteration months later. Does the file hosting actually cost peanuts when you do it yourself and the cloud has shattered my understanding of what it actually costs to deliver so much data?
At the scale of Huggingface, that still amounts to a lot load of money. Significantly less than if you did the same in AWS, but still a lot
That said, they do have a deal with AWS to make the data available in AWS ip space. Maybe they got some cheap hosting out of that too
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#178> For the Code Agent tasks among the public benchmarks above, DeepSeek-V4-Flash-0731 is evaluated with the minimal mode of DeepSeek Harness (to be released) as the agent framework So, are they planning to announce an optimized coding agent harness as well ? DSv4 flash is a fantastic model, and my daily driver. With reasonix or pi, I can code all day long and pay a few pennies for it. No token anxiety. Whereas the sam…
"For the Code Agent tasks among the public benchmarks above, DeepSeek-V4-Flash-0731 is evaluated with the minimal mode of DeepSeek Harness (to be released) as the agent framework, using the max reasoning effort level with temperature = 1.0, top_p = 0.95."
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#179Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#180https://files.parasmittal.com/openai_aa_luna_dsflash.svg
1: https://openai.com/index/advancing-the-price-performance-fro...