Live data from Hacker News

DeepSeek V4 Flash 0731

arcprize.org

481–490 of 507 posts

Re: DeepSeek V4 Flash 0731

#481
Well, I'm impressed.

Someone else here said we could get the model via OpenCode Go for $10/mo and get about $120 worth of credit, so I decided to give it a whirl. My first month is actually $5.

In the 3 hours I've been using it I've burned 3% of my 5-hr, 1% of my weekly and 0% of my monthly.

It's fixed 3 or 4 issues in my C++ game, despite not having visual capabilities to see the screenshots I was trying to give it. One-shotted them too.

Luna struggled with what I thought was an easy task (had to replace a few ASCII chars with the correct unicode char but kept choosing incorrectly).

I'll keep using it.

Re: DeepSeek V4 Flash 0731

#482

Earlier quoted context omitted.

To put actual numbers on it, since using AI to start solving all kinds of bottlenecks/inefficiencies in our small business, we've seen monthly net profit go up by around $4,000 USD. These are semi-permanent fixes, and the tech is only partially deployed. I am the only one using it, and I only use it part time. We've just spun up our first Hermes agent, with direct API access to our main inventory system and that's ex…

"оur first Hermes agent, with direct API access to our main inventory system" – let me assure you that absolutely nothing can go wrong here, mate. /s

An agent will read your comment and take it as a challenge to show something can go wrong. Thanks for prompt injecting, mate. /s

Re: DeepSeek V4 Flash 0731

#483

Earlier quoted context omitted.

That doesn’t make sense. It’s not like SOTA models are error free, yet we still use them. You use Fable 5 right? If that’s good enough for you now, why wouldn’t a Chinese model that’s as good as Fable 5 but at 10% the cost be good enough in 6 months?

"isn't the old version good enough ?" You can say this about literally every product we buy. And yet...

>"You can say this about literally every product we buy..." - and this is exactly why I am not buying many new things just for the fuck of it

Re: DeepSeek V4 Flash 0731

#484
post #162

Earlier quoted context omitted.

If what you're saying is true and accurate, then US-based AI labs are in big trouble. The only saving grace might be some sort of a 'national security' proclamation banning the use of state-of-the-art Chinese (and non-US) models across US federal and state governments and large enterprises (especially ones with federal government contracts), but even still, US AI labs will probably lose out massively on international…

One random says they are using DeepSeek and you proclaim it's all over for the top AI labs. Brilliant. First of all, no one knows the "true cost" of any of this, yet, but we know it's expensive. To what extent are the Chinese labs being subsidized? Are they real businesses? Second, the Chinese labs aren't some "super geniuses", while the American labs are full of clowns. As of today, like the past 3 years, American l…

>"To what extent are the Chinese labs being subsidized? Are they real businesses?"

And who gives a flying fuck. I am a "real" business and I count my money. It is not my life goal to prop some fat cats crying crocodile tears. Granted I do not use Chinese models. I use Junie straight from my JetBrain's IDEs that in turn uses Gemini Flash. Very cheap and more than enough for my use.

Re: DeepSeek V4 Flash 0731

#485
post #4

Kimi K3 was an interesting model only a month ago, and now we're looking at the same performance for 1/20th of the price. Wild how fast this is advancing.

Real question: is there anybody that is both maintaining alpha-dev capability by keeping abreast of all these daily changes, while also reserving enough time to actually work? Seems like we've reached the event horizon of whether AI advances are worth paying attention to.

I feel like it's most useful to get in a bit of a groove with one setup and then poke your head up every now and again and try updating. Staying at the bleeding edge of anything can be a bit of a treadmill.

Re: DeepSeek V4 Flash 0731

#486

Earlier quoted context omitted.

Well, yes, because the courts and the AI labs are totally separate entities. If they are doing it to you, they are probably doing it to others, which makes an easy class action

Is any class action lawsuit "easy"?

They lawyers do all the leg work (and keep most of the money), but from the defendant POV it all looks the same.

Re: DeepSeek V4 Flash 0731

#487

Earlier quoted context omitted.

When you do things in life, sometimes things go wrong. Oh no! This is such a ridiculous objection for how beloved it is. Wide swathes of the public can't cope with any adversity or risk.

No, the point is that with human in the loop the downside is (usually!) rather limited, as common sense would stop obvious fuckups (ok, not always, but still). With an agent (especially incompetently employed), the danger of unwittingly destroying your company (or at least, the crucial data/reputation) is rather higher. We are notoriously bad at estimating the downside risks in complex systems. The most obvious case…

There is a tradeoff to be considered between the utility gained from using the stuff, minus the risk severity/likelihood, plus available mitigations. As someone else posted, none of us are in a position to make that balanced post because we don't know if the guy has an airgapped backup or not etc.

In any case, GP's post was not such a balanced consideration; it was just parroting a beloved risk-aversion meme that can easily be deployed against building anything (what if the building falls on top of someone?) or even leaving home to go to work ("travelling in a hunk of steel at lethal speeds – let me assure you that absolutely nothing can go wrong here, mate.")

What I find tiresome about that meme is the presumption that "something can go wrong" is useful input on its own. It's not. Mistakes are made all the time, the only way to avoid that is to stop breathing. Even in the process of me standing up and going to the loo, something can go wrong.

If the guy wants to make a case that it's too dangerous for the expected benefits, he has to actually make that case. Saying "risk exists" with no elaboration is a waste of HTML. "something can go wrong" every time he swallows food, yet mysteriously he still does it.

(the suicide analogies may seem mean-spirited, but I kind of mean it. If you consider every action primarily from a standpoint of "what harm or irreversible change can result from this", the only permissible path is to do nothing. To be moral is to be as close as possible to a rock or another inanimate object.)

Re: DeepSeek V4 Flash 0731

#488

Earlier quoted context omitted.

Why? It's open weight, there are plenty providers on open router that are serving the latest v4 flash at 0.14/0.28 $.

This would be more convincing if those providers had converged on a number that was not the exact pricing of DeepSeek themselves. Clearly DeepSeek is setting the price here and without them holding it down I expect increases.

$0.14 is CN¥1. that's where the 0.14 comes from

Re: DeepSeek V4 Flash 0731

#489
post #464

Earlier quoted context omitted.

"оur first Hermes agent, with direct API access to our main inventory system" – let me assure you that absolutely nothing can go wrong here, mate. /s

Did you mean to post this on Reddit instead of HN?

I get what you were trying to say, but the quality difference is actually extremely low.
Post reply on HN