Live data from Hacker News

DeepSeek v4

api-docs.deepseek.com

991–1000 of 1001 posts

Re: DeepSeek v4

#991
The technical details in the paper are impressive, especially around the MoE architecture. It makes me think about the broader impact of these increasingly powerful models.

Re: DeepSeek v4

#992
post #590

somehow i canot open the link. but in their chinese version's release article, in the end ,there is a quote from xunzi( https://en.wikipedia.org/wiki/Xunzi_(philosopher) ) "Not seduced by praise, not terrified by slander; following the Way in one's conduct, and rectifying oneself with dignity." (不诱于誉,不恐于诽,率道而行,端然正己) (It is mainly used to express the way a Confucian gentleman conducts himself in the world. It reminds…

Sounds a lot like taoism, but i guess there's overlap

yeah..

Re: DeepSeek v4

#993

This is shockingly cheap for a near frontier model. This is insane. For context, for an agent we're working on, we're using 5-mini, which is $2/1m tokens. This is $0.30/1m tokens. And it's Opus 4.6 level - this can't be real. I am uncomfortable about sending user data which may contain PII to their servers in China so I won't be using this as appealing as it sounds. I need this to come to a US-hosted environment at a…

OpenAI released a privacy filter model https://openai.com/index/introducing-openai-privacy-filter . It is small enough to run locally, I would pass my request through it first before sending to any apis.

Re: DeepSeek v4

#994

They still don’t support json schema or batch api. It’s like deepseek does not want to make money

What do you currently use for json and batch, I was doing some analysis and my results show that gpt-oss-120b (non batch via openrotuer) is the best for now for my use case, better than gemini-flash models (batch on google). How is your experience?

Everything I do is json and of course you want that json in a specific format so that you can process it further.

Re: DeepSeek v4

#995

Earlier quoted context omitted.

American companies want a scan of your asshole for the privilege of paying to access their models, and unapologetically admit to storing, analyzing, training on, and freely giving your data to any authorities if requested. Chinese ulteriority is hypothetical, American is blatant.

I, personally, have never been asked for an asshole scan, but I'm interested in providing one if you can point me to a company that's offering.

rustysheriffsbadge.com

Re: DeepSeek v4

#996

Earlier quoted context omitted.

As a Brit I'm here for it to be honest, I'm tired of America with everything that's going on. China is not perfect but a bit of competition is healthy and needed

Americans are also tired with what’s going on.

Sure, but to claim that China is the better alternative is absurd.

Re: DeepSeek v4

#997
post #96

Earlier quoted context omitted.

Funny how Gemini is theoretically the best -- but in practice all the bugs in the interface mean I don't want to use it anymore. The worst is it forgets context (and lies about it), but it's very unreliable at reading pdfs (and lies about it). There's also no branch, so once the context is lost/polluted, you have to start projects over and build up the context from scratch again.

You know, with a bit of prompting, you can instruct Gemini to output the state of the conversation into a prompt that you can enter in a new chat and continue where you left off. But now with a fresh context window.

Not if Gemini Lost all context already. Also, it doesn't really work well, a lot of the nuance and information simply gets lost.
Post reply on HN