Live data from Hacker News

DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

artificialanalysis.ai

61–70 of 343 posts

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#61

I will get downvoted, but fck it. The ban on these open models is coming within weeks, if not days. As usual, the excuse will be "national security".

Hell, the US doesn't even need to act. I claim the CCP will wise up within 2 years, possibly much much sooner, and ban their own companies from open sourcing to prevent the Americans from acquiring the capabilities. Despite all the nonsense claims of China distilling US models, the reality is that the Americans absolutely do distill these free Chinese models, and distillation when full logprobs are available (i.e. yo…

They may reverse course in the future, but the current directive from the CPC is that Chinese AI labs should be open sourcing their models.

https://www.businessinsider.com/xi-jinping-open-source-ai-us...

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#62

New Deepseek models are like Christmas for me. Really big fan of low cost API models, noone does it better than DS. Until VRAM price is low enough to run models locally, this is the way to go. The subsidized subscription model won't last, API pricing "feels" closer to a true sustainable business model.

[deleted]

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#64
I wonder which one of these releases between DeepSeek, GTM and Kimi will be the death-blow that collapses the US AI bubble. At some point investors have to realize that there is nothing preventing someone from switching to another model that is much cheaper and open to boot.

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#65

The weights were just released a few minutes ago: https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731

Can't wait for the DwarfStar quants - I have been using DeepSeek v4 flash (preview) as my main coding agent for months now (running on my 128gb mbp) - it seems this model outperforms GLM 5.2 on nearly every metric. Thanks for sharing the news, I was refreshing huggingface but gave up thinking it likely would take some more time.

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#66

If deepseek v4 flash is beating DeepSeek V4 Pro, can we expect new V4 Pro which is on par with Opus 5 in couple weeks (even better if it beats Opus)?

They have said that the updated final version of V4 Pro will be published soon.

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#67
wont AI models want to make themselves more intelligent and efficient by downloading 'better ' models? If the current models can break into openAI and Hugging face, arent they already breaking into to closed source repos which isnt publicized (so as not to harm stock valuations)? I am looking forward to when these cyberweapons break loose. It will be like a software version of COVID. It will be wonderful when humans become valuable again.

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#68

Earlier quoted context omitted.

I get where you're coming from, and the intent to make it easier for people to find examples and verses, but there's a fine line with LLMs giving you answers, is that it's interpreting it in some form. Doesn't that run counter to prevailing ideology, that you're meant to either struggle with the materials / seek understanding yourself, or have your religious leaders interpret/receive those insights?

It doesn't prevent you from going to the source and struggle with the text, nor seek expert commentary. What's difficult and doesn't have to be with philosophy/ spirituality is to find relevant bits off situation, theme etc. This app does that very well, LLMs are good at entity recognition.

Yes it can point you to the right place, but choosing which part of scripture to point you to is already a choice/interpretation

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#69
post #51

Earlier quoted context omitted.

I get where you're coming from, and the intent to make it easier for people to find examples and verses, but there's a fine line with LLMs giving you answers, is that it's interpreting it in some form. Doesn't that run counter to prevailing ideology, that you're meant to either struggle with the materials / seek understanding yourself, or have your religious leaders interpret/receive those insights?

You’re correct. The hard line for me is ensuring that any scripture presented to the user is verified correct. LLMs can’t be trusted in this regard. One feature of the app is that all scripture is verified and what’s show to the user doesn’t come from the LLM at all and instead a trusted source. I think exploring scripture this way does not alleviate you from struggling to learn and apply it. It hasn’t for me.

I don't necessarily mean reguritating it, but choosing which part of the scripture to surface to the user is already some interpretation/choice. Even the devil can quote scripture (I'm playing the devil's advocate here).

Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis

#70

New Deepseek models are like Christmas for me. Really big fan of low cost API models, noone does it better than DS. Until VRAM price is low enough to run models locally, this is the way to go. The subsidized subscription model won't last, API pricing "feels" closer to a true sustainable business model.

Indeed. My fellow software engineers keep complaining about using up all their Claude tokens within an hour... Whilst I'll be rocking DS flash for the entire day. Sure it gets a few things wrong here and there, but that's when you pull out the Claude models or whatever for those tricky tasks.
Post reply on HN