Live data from Hacker News

DeepSeek v4

api-docs.deepseek.com

591–600 of 1001 posts

Re: DeepSeek v4

#591
post #110

Oh well, I should have bought 2x 512GB RAM MacStudios, not just one :(

Unironically curious about the performance of this model on unified VRAM machines.

Re: DeepSeek v4

#592

Just tested it via openrounter in the Pi Coding agent and it regularly fails to use the read and write tool correctly, very disappointing. Anyone know a fix besides prompting "always use the provided tools instead of writing your own call"

[deleted]

Re: DeepSeek v4

#593
post #426

Earlier quoted context omitted.

Still not sure how I feel about China of all places to control the only alternative AI stack, but I guess it's better than leaving everything to the US alone. If China ever feels emboldened enough to go for Taiwan and the US descends into complete chaos, the rest of the world running on AI will be at the mercy of authoritarian regimes. At the very least you can be sure noone is in this for the good of the people anym…

Isn’t Mistral close in the ballpark?

Mistral has a different focus. They aren't taking on trillions in debt risking their entire economy to produce useful products.

I think they are leaders in the democratization of LLMs. Almost everyone has a computer right now that can run a useful variant of a Mistral model. I hope they keep their focus because what they are aiming for likely has the biggest impact on the average person and would be the best case scenario for the technology in general.

Re: DeepSeek v4

#594
post #322

The incredible arrogance and hybris of the American initiated tech war - it is just a beautiful thing to see it slowly fall apart. The US-China contest aside - it is in the application layer llms will show their value. There the field, with llm commoditization and no clear monopolies, is wide open. There was a point in time where it looked like llms would the domain of a single well guarded monopoly - that would have…

I've been baffled watching America double down on the same strategy even when it failed to produce results They sanctioned the hell out of Huawei and now Huawei is bigger than ever America is just not able to digest the idea that another country can be as good, if not better, at innovation

Because it worked on Japan in the 80s and 90s and sometimes “Americans” have a hard time telling the two cultures apart.

Re: DeepSeek v4

#595

Open Source as it gets in this space, top notch developer documentation, and prices insanely low, while delivering frontier model capabilities. So basically, this is from hackers to hackers. Loving it! Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. In other words, Chinese ecosystem has delivered a complete AI stack. Like it or not, that's a big news. But what's there not to like when…

As a Brit I'm here for it to be honest, I'm tired of America with everything that's going on. China is not perfect but a bit of competition is healthy and needed

It’s a shame your country couldn’t get back its technical edge.

Re: DeepSeek v4

#596
post #322

The incredible arrogance and hybris of the American initiated tech war - it is just a beautiful thing to see it slowly fall apart. The US-China contest aside - it is in the application layer llms will show their value. There the field, with llm commoditization and no clear monopolies, is wide open. There was a point in time where it looked like llms would the domain of a single well guarded monopoly - that would have…

I've been baffled watching America double down on the same strategy even when it failed to produce results They sanctioned the hell out of Huawei and now Huawei is bigger than ever America is just not able to digest the idea that another country can be as good, if not better, at innovation

America has been making short term and short sighted moves to try to widen a gap that cannot sustain. They have chosen the wrong strategy out of fear and greed. Cooperation is the right strategy. Isolationism will not work in the long term except for maybe the handful that drove it. The irony is that it's an anticompetitive and anticapitalist move to do what they have been doing, so it's not even on principal.

Re: DeepSeek v4

#597
Something is odd with this model, their blog posts shows REALLY good results, but in most other third-party benchmarks, people realize it's not really SOTA, even bellow Kimi K2.6 and GLM-5/5.1

In my tests too[0], it doesn't reach top 10. One issue, which they also mentioned in their post, is that they can't really serve well the model at the moment, so V4-Pro is heavily rate-limited and gives a lot of timeout errors when I try to test it. This shouldn't be an issue though, considering the model is open-source, but it makes it hard to accurately test at the moment.

[0]: https://aibenchy.com/compare/deepseek-deepseek-v4-flash-high...

Re: DeepSeek v4

#598

> pricing "Pro" $3.48 / 1M output tokens vs $4.40 I’d like somebody to explain to me how the endless comments of "bleeding edge labs are subsidizing the inference at an insane rate" make sense in light of a humongous model like v4 pro being $4 per 1M. I’d bet even the subscriptions are profitable, much less the API prices. edit: $1.74/M input $3.48/M output on OpenRouter

Prices are not just hard cost of inference. Training costs are not equal. Chinese labs have cheaper access to large data centers. I also suspect they operate far more efficiently than orgs like openAI.

Re: DeepSeek v4

#599

Open Source as it gets in this space, top notch developer documentation, and prices insanely low, while delivering frontier model capabilities. So basically, this is from hackers to hackers. Loving it! Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. In other words, Chinese ecosystem has delivered a complete AI stack. Like it or not, that's a big news. But what's there not to like when…

I sometimes wonder if there are any security risks with using Chinese LLMs. Is there?

I sometimes wonder is there are any security risks with using LLMs from the US.

Re: DeepSeek v4

#600

It's interesting that they mentioned in the release notes: "Limited by the capacity of high-end computational resources, the current throughput of the Pro model remains constrained. We expect its pricing to decrease significantly once the Ascend 950 has been deployed into production." https://api-docs.deepseek.com/zh-cn/news/news260424#api-%E8%...

Yup, I tried to benchmark it, but harder questions time out or get rate-limited...
Post reply on HN