Live data from Hacker News

DeepSeek V4 Flash 0731

arcprize.org

391–400 of 481 posts

Re: DeepSeek V4 Flash 0731

#391
post #162

I've been using it extensively since the release and the best summary I can give is that it's good enough to use it for (almost) everything and cheap enough that the cost are irrelevant. I'm running it in Oh My Pi with a second instance running as "advisor" and even with 5-6 active sessions (effectively 12 streams) I'm struggling to spend more than 5 bucks per day. OpenCode Go even has double limits temporarily so fo…

If what you're saying is true and accurate, then US-based AI labs are in big trouble. The only saving grace might be some sort of a 'national security' proclamation banning the use of state-of-the-art Chinese (and non-US) models across US federal and state governments and large enterprises (especially ones with federal government contracts), but even still, US AI labs will probably lose out massively on international…

> If what you're saying is true and accurate, then US-based AI labs are in big trouble.

As it is now clear by behaviour of companies and US govt, all these investments will be backstopped by US govt. No US AI company will go hungry, they are national champions.

Re: DeepSeek V4 Flash 0731

#392

Earlier quoted context omitted.

If that’s the case businesses would be seeing millions to billions of profit gain (or cost reduction) in the past 4 months as they went from Opus 4.6 to Fable 5. But that’s simply not the case. It’s very clear that vast majority of the business do not generate additional value from incremental intelligence gain from these models. There is a reason why Chinese open weight models are now popular even in American enterp…

I work for a FAANG, and have my own personal projects for which I use the Chinese models. and the big models do indeed save/make us a lot of money. The Chinese models are not good enough for anything other than pair programming, which is just a very last-gen way of using agents. And when the big US models get better we will move with them. Until we stop seeing returns there is no "good enough", I don't know why this…

I use Chinese models, even smaller local ones, for much more than pair programming. If we are talking about deepseek v4 flash, which is basically a frontier model, it is much more capable than the local models I run on my MacBook Pro. The only issue really is finding the right harness.

I do have a way of correcting through redundancy, though. If you are just vibe coding, you need to use the most capable model you can find and even then it might not be good enough.

Re: DeepSeek V4 Flash 0731

#393
post #162

Earlier quoted context omitted.

If what you're saying is true and accurate, then US-based AI labs are in big trouble. The only saving grace might be some sort of a 'national security' proclamation banning the use of state-of-the-art Chinese (and non-US) models across US federal and state governments and large enterprises (especially ones with federal government contracts), but even still, US AI labs will probably lose out massively on international…

> only saving grace might be some sort of a 'national security' proclamation banning the use of state-of-the-art Chinese (and non-US) models In what kind of sad and failed dystopia is this a "saving grace"? For whom?

> "saving grace"? For whom?

For Anthropic and OpenAI, presumably. And the rather large economic distortion field around them, that may or may not go very badly for all our retirement funds if those firms become insolvent...

Re: DeepSeek V4 Flash 0731

#394
post #340
post #228

Earlier quoted context omitted.

If programming in the US to become unconditionally 10x more expensive, then the exodus from the US is about to begin.

Only to to find out they end up in a much worse place

> Only to to find out they end up in a much worse place

Like Europe?

Re: DeepSeek V4 Flash 0731

#395
post #164

Note this is the 07/31 release of DSv4 flash and not the "preview" that they put out a couple months or so ago. I've been running this model locally for a week, and the preview version before that. This updated one feels like a whole tier up. It's very capable for debugging and analyzing documents/data I upload. The killer feature, IMO, is the speed. On 2x RTX Pro 6000 Blackwell, its ~8k tok/s prefill and ~250 tok/s…

What runtime are you using with the 2x RTX Pro 6000 Blackwell machine? I have the same setup and tried DSv4 Flash on vLLM and ran into a ton of kernel bugs that don't seem to have been fixed yet.

This also works great: https://github.com/antirez/ds4

Re: DeepSeek V4 Flash 0731

#396

Earlier quoted context omitted.

5 USD is at the "raw" API price. OpenCode currently offers 60 USD API credits at 10 USD per month (OpenCode Go) and have even doubled it temporarily as a promotion. Effectively you can get Deepseek for 1/12th the already ridiculous cheap API price.

Those are not at the same price. Opecode’s 60 USD of deepseek usage is charged at much higher rates than what deepseek themselves charge at.

> Opecode’s 60 USD of deepseek usage is charged at much higher rates than what deepseek themselves charge at

Per the open code zen pricing page[1], it appears that the token prices are the same, but their cache is 10x more expensive?

[1]: https://opencode.ai/docs/zen/#pricing

Re: DeepSeek V4 Flash 0731

#397
post #269

Earlier quoted context omitted.

It's been true for almost every business. "Cheap and good enough" usually trumps "excellent but expensive". Ikea, McDonald's, Ryanair, AliExpress, Aldi - these brands prove that catering to poor people is more profitable than catering to rich people simply because there are so many poor people that their collective spending power outweights the one of rich people.

> RyanAir https://en.wikipedia.org/wiki/Category:Defunct_low-cost_airl...

https://en.wikipedia.org/wiki/List_of_defunct_airlines_of_th...

https://en.wikipedia.org/wiki/List_of_defunct_airlines_of_th...

https://en.wikipedia.org/wiki/List_of_defunct_airlines_of_th...

https://en.wikipedia.org/wiki/List_of_defunct_airlines_of_th...

Seems more like a overall industry problem, not limited to low cost carriers

Re: DeepSeek V4 Flash 0731

#398

Earlier quoted context omitted.

A collection of 30k-250k apps? Like individual unique apps?

Sorry, a collection of apps whose size is between 30kb and 250kb.

I read that and thought "ah that's going to confuse people, but I can tell they mean 30k-250k loc", so thank you for the clarification.

Re: DeepSeek V4 Flash 0731

#399
For the last 3 months I've been using V4 Flash Free with Hermes through Opencode Zen both personally and at my company and I've been having a great experience so far. It's my go-to model for terminal work, managing my entire ubuntu server, Cloudpanel, managing static websites, doing SEO audits, network tests, DNS troubleshooting, e-mail deliverability troubleshooting...

Furthermore, in my company we are using MCPs for Google Ads (it manages our ads), Analytics, Search Console, Zoho CRM, Microsoft Clarity... We use it to crawl specific websites and send daily summaries to our sales team in MS Teams channel. We use it to send daily summaries on marketing statistics and analytics... All with a FREE model. We are rarely hitting any limits so far and in case we need more tokens - we use NOUS or openrouter to pick between Flash or Pro for specific tasks that require more churning.

AMA.

Re: DeepSeek V4 Flash 0731

#400

Earlier quoted context omitted.

To put actual numbers on it, since using AI to start solving all kinds of bottlenecks/inefficiencies in our small business, we've seen monthly net profit go up by around $4,000 USD. These are semi-permanent fixes, and the tech is only partially deployed. I am the only one using it, and I only use it part time. We've just spun up our first Hermes agent, with direct API access to our main inventory system and that's ex…

"оur first Hermes agent, with direct API access to our main inventory system" – let me assure you that absolutely nothing can go wrong here, mate. /s

When you do things in life, sometimes things go wrong. Oh no!

This is such a ridiculous objection for how beloved it is. Wide swathes of the public can't cope with any adversity or risk.

Post reply on HN