Earlier quoted context omitted.
Still not sure how I feel about China of all places to control the only alternative AI stack, but I guess it's better than leaving everything to the US alone. If China ever feels emboldened enough to go for Taiwan and the US descends into complete chaos, the rest of the world running on AI will be at the mercy of authoritarian regimes. At the very least you can be sure noone is in this for the good of the people anym…
China doesn't even care about Taiwan anymore, their saber-rattling about it is a convenient distraction while they quietly make it completely irrelevant in the next few years.
DeepSeek v4
561–570 of 1001 posts
Re: DeepSeek v4
#562For those who rely on open source models but don't want to stop using frontier models, how do you manage it? Do you pay any of the Chinese subscription plans? Do you pay the API directly? After GPT 5.5 release, however good it is, I am a bit tired of this price hiking and reduced quota every week. I am now unemployed and cannot afford more expensive plans for the moment.
For DeepSeek you can use their API and if you ran it constantly you'd still be under what OpenAI or Anthropic charge for a coding plan.
I'm on Max x5 plan and any of the 'good' models like Kimi 2.6, GLM, DeepSeek would have cost 3-5x in per-token billing for what I used on my Claude plan the last three months
So unless my Claude fudged the maths to make itself look better, seems like I'm getting a good deal
Re: DeepSeek v4
#563Open Source as it gets in this space, top notch developer documentation, and prices insanely low, while delivering frontier model capabilities. So basically, this is from hackers to hackers. Loving it! Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. In other words, Chinese ecosystem has delivered a complete AI stack. Like it or not, that's a big news. But what's there not to like when…
Sorry, but exactly where did you get the idea that DS V4 runs entirely on Huawei? I asked DS itself and it denied this. It says: 'Nvidia chips are absolutely used for DeepSeek V4. The reality is a pragmatic "both-and" strategy, not an "either-or."' And based on the DS V4 technical report ( https://huggingface.co/deepseek-ai/DeepSeek-V4-Pro/blob/main... ), it is mentioned that: We validated the fine-grained EP scheme…
Bro, seriously?
Re: DeepSeek v4
#564Earlier quoted context omitted.
I sometimes wonder if there are any security risks with using Chinese LLMs. Is there?
Backdooring software at scale. Spearphishing. Building reliance and exploiting it, through state subsidies, dumping, and market manipulation. Handicapping provision to the west for competitive advantage.
Tech ceos are going around talking about how they will rule over employees and they will be unable to work in the future except for intelligence tokens. What if China commoditizes that without spending nearly as much resources? Kind of makes the trillions of dollars invested in the US a literal joke.
Re: DeepSeek v4
#565So is this the first AI lab using MUON for their frontier model?
No, Muon was developed by Moonshot; they've been using it in their Kimi models since Kimi K2 in 2025.
Re: DeepSeek v4
#566Earlier quoted context omitted.
Is it honestly better than Opus 4.6 or just benchmaxxed? Have you done any coding with an agent harness using it? If its coding abilities are better than Claude Code with Opus 4.6 then I will definitely be switching to this model.
Their Chinese announcement says that, based on internal employee testing, it is not as good as Opus 4.6 Thinking, but is slightly better than Opus 4.6 without Thinking enabled.
Re: DeepSeek v4
#567Re: DeepSeek v4
#568Earlier quoted context omitted.
[flagged]
Our (western) economic model forces competing individual companies to be profitable quickly. China can ignore DeepSeek losing money, because they know developing DeepSeek will help China. Not every institution needs to be profitable.
Re: DeepSeek v4
#569Earlier quoted context omitted.
This price is high even because of the current shortage of inference cards available to DeepSeek; they claimed in their press release that once the Ascend 950 computing cards are launched in the second half of the year, the price of the Pro version will drop significantly
In six month deepseek won't be sota anymore und usage will be wayyyy down.
Re: DeepSeek v4
#570I just want to remind you that this is happening at the same time as Anthropic A/B tests removal of Code from Pro Plan, and as OpenAI releases gpt-5.5 2x more expensive than gpt-5.4...