So GLM 5.2/Gemini 3.6 level intelligence for $0.28/m output. And their updated Pro model coming soon.... Plus a size you can genuinely run at home: Unsloth lossless Q8 at 162GB.
DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
131–140 of 342 posts
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#132New Deepseek models are like Christmas for me. Really big fan of low cost API models, noone does it better than DS. Until VRAM price is low enough to run models locally, this is the way to go. The subsidized subscription model won't last, API pricing "feels" closer to a true sustainable business model.
I would bet that Deepseek API pricing is still more cost effective per token than the subscriptions. With the increase in quality Deepseek Flash just got (in my personal testing so far, it seems to have improved a lot at following instructions, and has become more proactive), there really isn’t anything that can match it in terms of cost effectiveness.
I've used it in some open source code though, and loved how fast it was.
My mind is changing on how valuable my code actually is though... it's the complete picture, how it's put together, the design, the UI, the attention to detail that's the real value.
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#133Earlier quoted context omitted.
Can't wait for the DwarfStar quants - I have been using DeepSeek v4 flash (preview) as my main coding agent for months now (running on my 128gb mbp) - it seems this model outperforms GLM 5.2 on nearly every metric. Thanks for sharing the news, I was refreshing huggingface but gave up thinking it likely would take some more time.
What kind of tps are you getting?
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#134Earlier quoted context omitted.
[flagged]
I'm neither pro China, nor pro US. I'm pro open weights models, and I'm pro cheaper hardware. At this point I don't see any american frontier labs releasing SOTA open weights model, and I don't see ASML/Nvidia/Samsung monopoly getting any competition from anywhere apart from China in the near future.
yea i got that from your first comment ( although you removed crush American companies in _price_ ). you are pro cheapness at any cost even if its from your country's state funded direct geopolitical enemy.
China can always count on first order greed to win
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#135New Deepseek models are like Christmas for me. Really big fan of low cost API models, noone does it better than DS. Until VRAM price is low enough to run models locally, this is the way to go. The subsidized subscription model won't last, API pricing "feels" closer to a true sustainable business model.
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#136Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#137[flagged]
Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#138Re: DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
#139Daily reminder that none of these numbers are valid in a world where no one publishes the sampling settings used. Daily reminder that improving your samplers from the garbage default top_p/top_k to min_p or subsequent methods dramatically improves the performance of these models, and makes most quantities like measured "verbosity" and subsequent calculations of "intelligence per token" meaningless Daily reminder that…