Live data from Hacker News

DeepSeek v4

api-docs.deepseek.com

831–840 of 1001 posts

Re: DeepSeek v4

#831

Earlier quoted context omitted.

>- it is just a beautiful thing to see it slowly fall apart. I feel uneasy over China dominance as much as the US. I trust US more still as Europe has a post WW2 relationship. I notice many comments being pro China but they seem to be from the third world (one mentioned a very low salary) I feel the opening of the internet was a mistake. China is a totilitarian dictatorship. This is a fact. Look into Mistral AI too :…

Honestly the China scare mongering is borderline hilarious. The US has literally attacked two countries this year in the span of weeks and is blockading another causing needless deaths. Not too mention the last 50 years of US imperialism making the world a worse place for everyone, except to benefit the few (doesn't benefit Americans, only capitalists). The idea that China is worse than America is laughable. LMK when…

One party authoritarian dictatorship with no free speech or democratic elections and no civil rights movement seems pretty bad to me. No amount of whataboutism is ever going to compete with that.

It also seems like clashes with India, every southeast asian country with internationally recognized territory rights in the South China sea, the forcible takeover of Hong Kong, arming and economically supporting Russia, Pakistan and Iran are bad, and the increasing probability of a hot war to take over Taiwan should count as bad, perhaps the most urgently dangerous threat to global peace in the 21st century.

The United States track record post WW2 is a complicated combination of monstrously immoral Kissenger and Bush style overthrows of democracies and genuinely valuable maintenance of a post WW2 democratic order focused on things like free speech and human rights. I stay with full sincerity that in the decade plus that I've been here on hn seeing whataboutism as a strategy for defending China, I'm yet to encounter anything that feels like a sincere engagement with United States role in the world as a combination of positives and negatives, it's always flatly one-sided messaging that feels like it's aimed at a favorable audience that already agree rather than like it's sincerely attempting to persuade.

Re: DeepSeek v4

#832

Earlier quoted context omitted.

On-device is incredibly far away from being viable. A $20 ChatGPT subscription beats the hell out of the 8B model that a $1,000 computer can run. Nvidia's forward PE ratio is only 20 for 2026. That's much lower than companies like Walmart and Costco. It's also growing nearly 100% YoY and has a $1 trillion backlog. I think Nvidia is cheap.

I think you overestimate what most people are doing with AI. A 2B model can give out relationship advice and tell you how long to boil an egg.

And honestly, what other types of questions would you ever need answers to?

Re: DeepSeek v4

#833
post #763

Earlier quoted context omitted.

>- it is just a beautiful thing to see it slowly fall apart. I feel uneasy over China dominance as much as the US. I trust US more still as Europe has a post WW2 relationship. I notice many comments being pro China but they seem to be from the third world (one mentioned a very low salary) I feel the opening of the internet was a mistake. China is a totilitarian dictatorship. This is a fact. Look into Mistral AI too :…

Are third world users opinions of lesser value?

Come on, Sweden isn’t quite a 3rd world country.

Re: DeepSeek v4

#834

Earlier quoted context omitted.

> The idea that for basically sub-agents, we can fine-tune them, should reasonably expect to perform as well as Opus for a specific subtask of which my applications have many [...] we can run a general-purpose intelligent model, Sonnet or Opus, orchestrating a fleet of, let's say, 30 to 50 of these sub-agents that have been fine-tuned I've heard so many people saying this for the last year, and even tried doing it my…

I guess it depends on a task. Opus is already spawning Sonnet/Haiku for simple tasks with a good success rate.

I think "agent spawns weaker agent to do safe edit sometimes" is vastly different than the imagined "general-purpose intelligent model orchestrating a fleet of 50 sub-agents".

Re: DeepSeek v4

#835

It's easy to praise Deepseek for its results and generosity -- how they can keep up with frontier labs on Huawei chips for a fraction of the cost! -- but let's not forget a big part of their toolkit is heavy distillation of SoTA.

So they distill the sota model where OAI/Anthropic illegally stole from public, and open weights to us or sell their API at 1/50th of the price? I'd say keep up the good work and distill more!

Re: DeepSeek v4

#836
post #827

Earlier quoted context omitted.

> Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. That is a huge claim to make with no evidence. I researched what you said, and I have found no statement to that effect in their paper[0], on huggingface[1], twitter[2], WeChat[3], or in their news release[4]. They only mention as a footnote in only the Chinese version of their news release that they plan to reduce inference costs with…

DeepSeek is planning to use Huawei extensively for inference “Due to constraints in high-end compute capacity, the current service capacity for Pro is very limited. After the 950 supernodes are launched at scale in the second half of this year, the price of Pro is expected to be reduced significantly.” https://x.com/jukan05/status/2047516566149816627

Yes, that's the footnote from citation [5].

Re: DeepSeek v4

#837
post #510

There are quite a few comments here about benchmark and coding performance. I would like to offer some opinions regarding its capacity for mathematics problems in an active research setting. I have a collection of novel probability and statistics problems at the masters and PhD level with varying degrees of feasibility. My test suite involves running these problems through first (often with about 2-6 papers for conte…

I reviewed how DeepSeek V4-Pro, Kimi 2.6, Opus 4.6, and Opus 4.7 across the same AI benchmarks. All results are for Max editions, except for Kimi. Summary: Opus 4.6 forms the baseline all three are trying to beat. DeepSeek V4-Pro roughly matches it across the board, Kimi K2.6 edges it on agentic/coding benchmarks, and Opus 4.7 surpasses it on nearly everything except web search. DeepSeek V4-Pro Max shines in competit…

I'd be interested to know when that Opus 4.6 baseline is from given their recent recognition of performance issues. Do you have a paper posted on this review?

Re: DeepSeek v4

#838

Objective, detailed benchmark results at https://gertlabs.com Early takeaways: from this release, DeepSeek V4 Flash is the model to pay attention to here. It's cheap, effective, and REALLY fast. The Pro model is slow, not much better in coding reasoning so far when it works, and honestly too unreliable and rate limited to be of much use, currently. Hopefully that improves as new providers host the model. Flash is wor…

I would say all benchmarks are inherently subjective. How is yours better? It seems to produce a little bit strange results. Opus 4.6 being worse than 4.5 for example. Or chinese models being rated too high. Kimi, Deepseek or GLM are all great in open source world, but I don't believe they are ahead of SOTA models from Anthropic, OpenAI or Google.

Re: DeepSeek v4

#839

Earlier quoted context omitted.

They have had the best math models for about a year most folks just didn't know about it. You can't find inference on APIs, but I run these at home, this is also the advantage of open models. https://huggingface.co/deepseek-ai/DeepSeek-Math-V2 https://huggingface.co/deepseek-ai/DeepSeek-Prover-V2-671B

You run a 671B model at home?

It's a big house.

Re: DeepSeek v4

#840

Earlier quoted context omitted.

They have had the best math models for about a year most folks just didn't know about it. You can't find inference on APIs, but I run these at home, this is also the advantage of open models. https://huggingface.co/deepseek-ai/DeepSeek-Math-V2 https://huggingface.co/deepseek-ai/DeepSeek-Prover-V2-671B

You run a 671B model at home?

[deleted]
Post reply on HN