Earlier quoted context omitted.
You're using it on low, that's why. There's a huge difference in performance from low to max effort.
I’m comparing same / similar settings between models. I can’t use high on one and low on others it’s not a fair test. Not sure why I was downvoted. But seems the downvoter is quick to downvote anything that doesn’t fit the narrative they’re looking for. I’m just reporting my findings.
DeepSeek V4 Flash 0731
361–370 of 481 posts
Re: DeepSeek V4 Flash 0731
#362Seeing everyone spend like 200USD a month seems kind of mad. I have £20/month Gemini and £20 a month claude for a bunch of personal projects. Yes I have to wait sometimes, it's probably a good thing.
Re: DeepSeek V4 Flash 0731
#363I've been using it extensively since the release and the best summary I can give is that it's good enough to use it for (almost) everything and cheap enough that the cost are irrelevant. I'm running it in Oh My Pi with a second instance running as "advisor" and even with 5-6 active sessions (effectively 12 streams) I'm struggling to spend more than 5 bucks per day. OpenCode Go even has double limits temporarily so fo…
If what you're saying is true and accurate, then US-based AI labs are in big trouble. The only saving grace might be some sort of a 'national security' proclamation banning the use of state-of-the-art Chinese (and non-US) models across US federal and state governments and large enterprises (especially ones with federal government contracts), but even still, US AI labs will probably lose out massively on international…
this has already happened with manufacturing, so it isn't surprising that other industries follow.
The US premium in engineering and scientific endeavors have been lacking for the past 30-40 years, and if it werent for tech and silicon valley, the US would have nothing state of the art. Even on that front, the US is falling behind given how much effort in tech has been diverted into privacy invading, and advertising.
The US has been riding momentum, but eventually that momentum will stop. It will take half a century to get back up to speed, and by then, the US will have fallen behind so far that catching back up will seem impossible.
The telling evidence would be if china has the first moonbase before the US does. I think this is highly likely looking at today's US administration.
Re: DeepSeek V4 Flash 0731
#364Seeing everyone spend like 200USD a month seems kind of mad. I have £20/month Gemini and £20 a month claude for a bunch of personal projects. Yes I have to wait sometimes, it's probably a good thing.
I max out my $200/month Claude plan. You are obviously just not taking advantage of it to the same level as others. Which is fine. Don't pay for something you don't need. But I would definitely take a massive productivity hit if I had 1/20th the usage.
Re: DeepSeek V4 Flash 0731
#365Earlier quoted context omitted.
If what you're saying is true and accurate, then US-based AI labs are in big trouble. The only saving grace might be some sort of a 'national security' proclamation banning the use of state-of-the-art Chinese (and non-US) models across US federal and state governments and large enterprises (especially ones with federal government contracts), but even still, US AI labs will probably lose out massively on international…
I think you're overlooking the fact that for long-horizon tasks, even small errors compound over time and can lead to catastrophic outcomes. For simple queries, we have reached the threshold since the beginning of the year, and models are good enough from every provider to make a meaningful difference between one another. (ChatGPT, Claude, Gemini, Grok, MuseSpark, Kimi, DeepSeek, GLM...) The real unlock will be, and…
Re: DeepSeek V4 Flash 0731
#366Earlier quoted context omitted.
I wonder when we crossed the "99 percentile of intelligence for 99% of the usecases" threshold. At this point, the gains seem to be right at the very edge of bleeding edge for narrow and specialized use cases, and wonder if it'll be a sort of diminishing return from here on.
Probably the best counterexample is the games they are able to design. It's still mostly AI slop, few would want to play.
Re: DeepSeek V4 Flash 0731
#367Earlier quoted context omitted.
Except Fable won’t be costing $100 for enterprises that will be considering the Chinese models. If $100 Claud Max subscription works for you, then great. But you have to remember your pricing is subsidized by enterprises that pay hundreds of thousands of dollars each month, if not more, to Anthropic. For those companies, a Chinese model that can cut their AI spend from $1M/month to $200k suddenly seems attractive. An…
I don’t think the industry knows how to price this stuff. Deepseek is great (I’m running it on a RTX 6000 pro setup) but it’s nothing like Fable. It’s still strongly human-in-the-loop which is fine, until you experience how good these models can be. Think about it this way. Let’s say you could buy an LLM that gets things right 98% of the time. But there’s another LLM that’s 100x the price but gets things right 99.9%…
But that’s simply not the case. It’s very clear that vast majority of the business do not generate additional value from incremental intelligence gain from these models.
There is a reason why Chinese open weight models are now popular even in American enterprises, because CTOs realize that they are indeed good enough.
Re: DeepSeek V4 Flash 0731
#368I've been using it extensively since the release and the best summary I can give is that it's good enough to use it for (almost) everything and cheap enough that the cost are irrelevant. I'm running it in Oh My Pi with a second instance running as "advisor" and even with 5-6 active sessions (effectively 12 streams) I'm struggling to spend more than 5 bucks per day. OpenCode Go even has double limits temporarily so fo…
If what you're saying is true and accurate, then US-based AI labs are in big trouble. The only saving grace might be some sort of a 'national security' proclamation banning the use of state-of-the-art Chinese (and non-US) models across US federal and state governments and large enterprises (especially ones with federal government contracts), but even still, US AI labs will probably lose out massively on international…
Re: DeepSeek V4 Flash 0731
#369Earlier quoted context omitted.
If what you're saying is true and accurate, then US-based AI labs are in big trouble. The only saving grace might be some sort of a 'national security' proclamation banning the use of state-of-the-art Chinese (and non-US) models across US federal and state governments and large enterprises (especially ones with federal government contracts), but even still, US AI labs will probably lose out massively on international…
I think you're overlooking the fact that for long-horizon tasks, even small errors compound over time and can lead to catastrophic outcomes. For simple queries, we have reached the threshold since the beginning of the year, and models are good enough from every provider to make a meaningful difference between one another. (ChatGPT, Claude, Gemini, Grok, MuseSpark, Kimi, DeepSeek, GLM...) The real unlock will be, and…
Re: DeepSeek V4 Flash 0731
#370Note this is the 07/31 release of DSv4 flash and not the "preview" that they put out a couple months or so ago. I've been running this model locally for a week, and the preview version before that. This updated one feels like a whole tier up. It's very capable for debugging and analyzing documents/data I upload. The killer feature, IMO, is the speed. On 2x RTX Pro 6000 Blackwell, its ~8k tok/s prefill and ~250 tok/s…
For reference, on a 1x B300 it's over 400 tok/s decode on a single stream.