Live data from Hacker News

DeepSeek v4

api-docs.deepseek.com

861–870 of 1001 posts

Re: DeepSeek v4

#861

Earlier quoted context omitted.

> Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. That is a huge claim to make with no evidence. I researched what you said, and I have found no statement to that effect in their paper[0], on huggingface[1], twitter[2], WeChat[3], or in their news release[4]. They only mention as a footnote in only the Chinese version of their news release that they plan to reduce inference costs with…

Here's a note about running entirely on Huawei chips: https://finance.yahoo.com/sectors/technology/articles/deepse...

> DeepSeek indicated that current service capacity for the V4 Pro series is constrained by a computing crunch, though pricing could fall after new clusters powered by Huawei's Ascend 950 chips come online in the second half of the year.

Only mention of Huawei in that article (as of now).

Re: DeepSeek v4

#862

Earlier quoted context omitted.

> Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. That is a huge claim to make with no evidence. I researched what you said, and I have found no statement to that effect in their paper[0], on huggingface[1], twitter[2], WeChat[3], or in their news release[4]. They only mention as a footnote in only the Chinese version of their news release that they plan to reduce inference costs with…

Here's a note about running entirely on Huawei chips: https://finance.yahoo.com/sectors/technology/articles/deepse...

Did you read any part of the link you posted? Huawei is mentioned once and not in the context of the model being trained or currently running on Huawei chips.

Re: DeepSeek v4

#863

Open Source as it gets in this space, top notch developer documentation, and prices insanely low, while delivering frontier model capabilities. So basically, this is from hackers to hackers. Loving it! Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. In other words, Chinese ecosystem has delivered a complete AI stack. Like it or not, that's a big news. But what's there not to like when…

I am all for monopoly breakdown. But there is an argument that this is anticompetitive strategy designed to undercut the commercial viability of the other labs. In free trade negotiations this is called “dumping”: selling a product below cost at a high volume to gain market share by driving competition out of the market and then raising prices when you’ve outlasted them.

Re: DeepSeek v4

#864

There are quite a few comments here about benchmark and coding performance. I would like to offer some opinions regarding its capacity for mathematics problems in an active research setting. I have a collection of novel probability and statistics problems at the masters and PhD level with varying degrees of feasibility. My test suite involves running these problems through first (often with about 2-6 papers for conte…

Any plans to publish the benchmark results?

Re: DeepSeek v4

#865

Earlier quoted context omitted.

You run a 671B model at home?

Yes, and plenty of others do too. Quantizied. Join us at r/localllama My largest models 318G /llmzoo/models/Qwen3.5-397B 377G DeepSeekv3.2-nolight 380G /llmzoo/models/DeepSeek-V3.2-UD 400G /llmzoo/models/Qwen3.5-397B-Q8 443G DeepSeek-Math-v2 443G DeepSeek-V3-0324-Q5 522G /llmzoo/models/GLM5.1 545G /llmzoo/models/kimi2.6 546G /llmzoo/models/KimiK2.5

What hardware do you use?

Re: DeepSeek v4

#866
post #670

Seriously, why can't huge companies like OpenAI and Google produce documentation that is half this good?? https://api-docs.deepseek.com/guides/thinking_mode No BS, just a concise description of exactly what I need to write my own agent.

I spent only two minutes reading their documentation and it’s clear no one did any proofreading and it’s full of mistakes made by non-native speakers. Example: the second sentence on the first page says “softwares” but “software” is a mass noun that cannot be pluralized. Example: the third page about tokens has some zipped code to “calculate the token usage for your intput/output” and obviously “intput” should be “in…

i prefer it cuz it indicates they didnt use an LLM to write their documentations and that its human generated

Re: DeepSeek v4

#867
post #806
post #670

Earlier quoted context omitted.

I spent only two minutes reading their documentation and it’s clear no one did any proofreading and it’s full of mistakes made by non-native speakers. Example: the second sentence on the first page says “softwares” but “software” is a mass noun that cannot be pluralized. Example: the third page about tokens has some zipped code to “calculate the token usage for your intput/output” and obviously “intput” should be “in…

No one cares about this kind of stuff. 99% of the devs are not English native speakers, what do you expect ? It works and we all can understand it

I try hard not to care but subconsciously spelling errors and grammar issues scream low-quality work to me. It’s the kind of mistake that’s the easiest to correct, and they didn’t bother.

Re: DeepSeek v4

#868
post #843

Just tested it via openrounter in the Pi Coding agent and it regularly fails to use the read and write tool correctly, very disappointing. Anyone know a fix besides prompting "always use the provided tools instead of writing your own call"

If you have access to any other model it can create create pi extension that fixes problem. At least worked for me.

Like a special parser? Would you mind elaborating?

Re: DeepSeek v4

#869
post #867
post #806

Earlier quoted context omitted.

No one cares about this kind of stuff. 99% of the devs are not English native speakers, what do you expect ? It works and we all can understand it

I try hard not to care but subconsciously spelling errors and grammar issues scream low-quality work to me. It’s the kind of mistake that’s the easiest to correct, and they didn’t bother.

That seems like a you problem

Re: DeepSeek v4

#870

> pricing "Pro" $3.48 / 1M output tokens vs $4.40 I’d like somebody to explain to me how the endless comments of "bleeding edge labs are subsidizing the inference at an insane rate" make sense in light of a humongous model like v4 pro being $4 per 1M. I’d bet even the subscriptions are profitable, much less the API prices. edit: $1.74/M input $3.48/M output on OpenRouter

Insert always has been meme. But seriously, it just stems from the fact some people want AI to go away. If you set your conclusion first, you can very easily derive any premise. AI must go away -> AI must be a bad business -> AI must be losing money.

[deleted]
Post reply on HN