Live data from Hacker News

DeepSeek v4

api-docs.deepseek.com

571–580 of 1001 posts

Re: DeepSeek v4

#571
post #355

Earlier quoted context omitted.

I don’t know if we’re ahead of the curve but that tired feeling has started turning into hate here in the EU. I guess being threatened with invasion does that to you. The next decade is going to look very different with America Alone.

I grew up in the states when I was younger, always feeling some closeness to Americans even after I moved back to Europe. With all that goes on it has changed. Recently I sat on a plane near some Americans discussing their holidays here, and I noticed I felt contempt. Sitting their with insane privilege as their government torches the world. Individuals remain individuals, and one really ought not to be prejudice. Ho…

>However the lack of resistance I see in in the “land of the free” as their “democratic” institutions collapse just makes me believe they never cared at all.

Largest protests in US history just in the past year:

https://en.wikipedia.org/wiki/List_of_protests_and_demonstra...

>insane privilege

My sister and brother recently graduated from college, have been searching for jobs for over 6 months, they can't find anything. They're politically liberal Californians.

Re: DeepSeek v4

#572

Assuming it is almost as good as Opus 4.6 (which benchmarks seem to give evidence for), and assuming we are having a good enough harness (PI, OpenCode), it's is now more than 5x cheaper. I just want to remind you that this is happening at the same time as Anthropic A/B tests removal of Code from Pro Plan, and as OpenAI releases gpt-5.5 2x more expensive than gpt-5.4...

> Assuming it is almost as good as Opus 4.6 (which benchmarks seem to give evidence for)

That’s a big if. It’s my experience that models that perform very well on benchmarks do not necessarily perform well in real life.

I’ve mostly started ignoring the benchmarks and run my own evals.

Re: DeepSeek v4

#573

Earlier quoted context omitted.

it's in the footnote text of the first figure of the section the link points to, where "昇腾950" means "Ascend 950"

OK, strange that it doesn't appear on my version of the webpage https://api-docs.deepseek.com/zh-cn/news/news260424#api-%E8%... This is the first figure of the section that the above links point to ( https://api-docs.deepseek.com/zh-cn/img/v4-spec.png ). And I can read Chinese.

https://api-docs.deepseek.com/zh-cn/img/v4-price.png

Re: DeepSeek v4

#574

Earlier quoted context omitted.

Can't see how NVIDA justifies its valuation/forward P/E ratio with these developments and on-device also becoming viable for 98% of people's needs when it comes to AI

On-device is incredibly far away from being viable. A $20 ChatGPT subscription beats the hell out of the 8B model that a $1,000 computer can run. Nvidia's forward PE ratio is only 20 for 2026. That's much lower than companies like Walmart and Costco. It's also growing nearly 100% YoY and has a $1 trillion backlog. I think Nvidia is cheap.

8b models can run on laptops. Of course a 1.8T model is more capable, but for a lot of tasks it really isn't 1000x

Re: DeepSeek v4

#575
post #42

History doesn't always repeat itself. But if it does, then in the following week we'll see DeepSeek4 floods every AI-related online space. Thousands of posts swearing how it's better than the latest models OpenAI/Anthropic/Google have but only costs pennies. Then a few weeks later it'll be forgotten by most.

It's difficult because even if the underlying model is very good, not having a pre-built harness like Claude Code makes it very un-sticky for most devs. Even at equal quality, the friction (or at least perceived friction) is higher than the mainstream models.

You can literally run it from Claude code. Easily too

Re: DeepSeek v4

#576

For those who rely on open source models but don't want to stop using frontier models, how do you manage it? Do you pay any of the Chinese subscription plans? Do you pay the API directly? After GPT 5.5 release, however good it is, I am a bit tired of this price hiking and reduced quota every week. I am now unemployed and cannot afford more expensive plans for the moment.

At home I currently use MiniMax via OpenRouter - it’s pretty good and very cheap. They have a subscription plan, but I’m not ready to commit to it yet. Another way to keep the ability to try out new models is to buy a reseller subscription like Cursor’s.

I tried OpenRouter but I feel the money flies even with these models, it is not comparable to a subscription but yes, it's very good for trying. Maybe I should test other models alongside GPT 5.5 to see which one fits me.

Re: DeepSeek v4

#577

For those who rely on open source models but don't want to stop using frontier models, how do you manage it? Do you pay any of the Chinese subscription plans? Do you pay the API directly? After GPT 5.5 release, however good it is, I am a bit tired of this price hiking and reduced quota every week. I am now unemployed and cannot afford more expensive plans for the moment.

I've been on Kimi K2.5 on openrouter for a couple of months for anything I can't run locally. Really is dirt cheap for how good it is. Haven't assessed K2.6 yet but the price is higher so it needs to be more efficient, not just more capable.

But more broadly: openrouter solves the problem of making a broad range of models available with a single payment endpoint, so you can just switch around as much as you like.

Re: DeepSeek v4

#578

For those who rely on open source models but don't want to stop using frontier models, how do you manage it? Do you pay any of the Chinese subscription plans? Do you pay the API directly? After GPT 5.5 release, however good it is, I am a bit tired of this price hiking and reduced quota every week. I am now unemployed and cannot afford more expensive plans for the moment.

For DeepSeek you can use their API and if you ran it constantly you'd still be under what OpenAI or Anthropic charge for a coding plan.

I am not so sure, credits fly when using any model trough API if I use it as much as I use Codex.

Re: DeepSeek v4

#579
post #29

The Flash version is 284B A13B in mixed FP8 / FP4 and the full native precision weights total approximately 154 GB. KV cache is said to take 10% as much space as V3. This looks very accessible for people running "large" local models. It's a nice follow up to the Gemma 4 and Qwen3.5 small local models.

I'm going to blow my bandwidth allowance again this month, aren't I.

Re: DeepSeek v4

#580

Open Source as it gets in this space, top notch developer documentation, and prices insanely low, while delivering frontier model capabilities. So basically, this is from hackers to hackers. Loving it! Also, note that there's zero CUDA dependency. It runs entirely on Huawei chips. In other words, Chinese ecosystem has delivered a complete AI stack. Like it or not, that's a big news. But what's there not to like when…

As a Chinese, I feel tiered, it's like the cold war, what is takes to keep competitive with every aspect, it's just another win for the country and the corp
Post reply on HN