Live data from Hacker News

Notes on DeepSeek

news.ycombinator.com

41–50 of 157 posts

Re: Notes on DeepSeek

#41

Post appears to have been removed, I caught a copy of it: https://pastebin.com/rcAqEFG1 I assume it will get reposted at some point.

thanks. this really isnt that long, might as well paste in full here since OP deleted.

Notes on DeepSeek:

We visited the company HQ last Tuesday. It was founded in 2023 by Liang Wenfeng and operated out of his hedge fund, High-Flyer, until somewhat recently. The company released their R1 model in January 2025, so it was interesting to see what they’ve been doing

The company is located in an unmarked, 12-story building in Hangzhou. There is no DeepSeek branding visible from the street or lobby. I asked why this is, and the team demurred and said, “Well, there are many companies in this building, and we are not special.” They want to keep a low profile.

We met with their Head of Data and Head of Infrastructure. The company only has 300 employees. They are at least an order-of-magnitude smaller than Anthropic, and don’t care to scale further just yet. Their Head of Infrastructure, in particular, was young; maybe 30 years old and apparently one of the best AI buildout and energy experts in the country. (We briefly walked through the labs, and everybody seemed young. There was a lot of discussion; it felt like an exciting and energetic place.)

Lots of competition is coming from Alibaba (Qwen), ByteDance, and Moonshot (Kimi). People in China seem to mostly use Kimi or Deepseek. Young people use VPNs to access Claude, though Anthropic has blockers around usage in China and make it difficult. Poaching between groups is common, just like in the U.S. DeepSeek has a reputation as being really smart and “cool,” maybe similar to Anthropic. Big labs are mostly in Beijing, near Tsinghua and Peking University, with Hangzhou as the main exception (DeepSeek and Alibaba/Qwen are there).

The DeepSeek team reads western AI writers. They listen to Dwarkesh and read Gwern. The people we met with said they had never met with any employees from Anthropic. They were not at all concerned with some kind of hostile / AGI takeover scenario. They kept bringing up job loss (which is already high amongst youth in China) as their main concern. When we asked if they do red teaming on their models, they said no. In China, AI models are not regulated directly; the government instead has restrictions on how those models can be used in software, services, etc.

As a whole, China seems to treat AI as just another technology, rather than as some kind of singularity moment. National attention is still on basic needs and infrastructure buildouts, and on providing more medicines for people. The “dreams of singularity" seem like a luxury or distant consideration.

We asked the DeepSeek team: “What has the highlight been so far? What are your plans for an exit?” And they said that their highlight and great achievement was R1. They did not gesticulate at a future model or vision, but rather seemed proudest of what they’ve already done. They are content for now to remain ~6 months behind U.S. companies while maintaining a lower profile and team size.

Re: Notes on DeepSeek

#43
post #41

Post appears to have been removed, I caught a copy of it: https://pastebin.com/rcAqEFG1 I assume it will get reposted at some point.

thanks. this really isnt that long, might as well paste in full here since OP deleted. Notes on DeepSeek: We visited the company HQ last Tuesday. It was founded in 2023 by Liang Wenfeng and operated out of his hedge fund, High-Flyer, until somewhat recently. The company released their R1 model in January 2025, so it was interesting to see what they’ve been doing The company is located in an unmarked, 12-story buildin…

[deleted]

Re: Notes on DeepSeek

#44
post #12

Not sure what I read, but sounded like a lunch meeting description; felt void of actual information, with the restaurant replaced by the office. I am in China and can tell it is either Kimi, DeepSeek or Claude (proxied or actually deepseek/fake). The bigger push for the general public died down a lot since last year; kids were pushed to use AI for homework, now it is disallowed and frowned upon. In short mixed messag…

Were things like "300 employees" and descriptions of the deliberately low key hdq out there before? That counts as actual information to me.

Re: Notes on DeepSeek

#45

Earlier quoted context omitted.

Yep. After yesterday's moves around "Fable 5" even twice as much. We've had a taste, and damned if I'm going to have the "means of production" snatched from me already?

approximately how many months/years until there are "illegal models"?

You wouldn’t steal a brain

Re: Notes on DeepSeek

#46
post #41

Post appears to have been removed, I caught a copy of it: https://pastebin.com/rcAqEFG1 I assume it will get reposted at some point.

thanks. this really isnt that long, might as well paste in full here since OP deleted. Notes on DeepSeek: We visited the company HQ last Tuesday. It was founded in 2023 by Liang Wenfeng and operated out of his hedge fund, High-Flyer, until somewhat recently. The company released their R1 model in January 2025, so it was interesting to see what they’ve been doing The company is located in an unmarked, 12-story buildin…

I don't get the part of "AI models are not regulated directly, the government instead has restrictions on how those models can be used in software, services". Is it not the same thing? When I chat with DeepSeek about any (Chinese) political/social issue, it immediately begins aligning with the party's line or just cut off the conversation abruptly.

Re: Notes on DeepSeek

#47
post #26

Earlier quoted context omitted.

There are 2 things worth separating. 1) China distills and is therefore morally bad. As you rightly point out, that's not a great argument. 2) China distills and is therefore possibly not that competent. I think that makes sense. If they only catch up to the frontier through distillation then 1) Their model will never be as good as the model they are distilling from. 2) They will never reach the frontier - they need…

This is literally a repeat of the whole “China only make low quality cheap stuff” argument. “All they do is copy.” And now, oops they are world leaders in EVs, batteries, solar, drones, just to name a few on the biggest consumer facing things.

"Success leaves clues"

You gotta start somewhere and you can start at page 1 or page 10 and that time, energy and cost you saved starting 9 pages later can be put into making whatever it is you're building better than the original.

The US, and every other country, is full of derivatives or straight up copies. No one is getting super mad at the generic cheerios at the grocery store. It's hypocrisy.

Re: Notes on DeepSeek

#48
post #12

Not sure what I read, but sounded like a lunch meeting description; felt void of actual information, with the restaurant replaced by the office. I am in China and can tell it is either Kimi, DeepSeek or Claude (proxied or actually deepseek/fake). The bigger push for the general public died down a lot since last year; kids were pushed to use AI for homework, now it is disallowed and frowned upon. In short mixed messag…

With government billions fund pushed for AI build out, fast pace integration on large scale and sweeping national education reform for AI, I don't think it can be called "died down".

[0] https://www.reuters.com/world/china/china-prepares-295-billi...

[1] https://www.globalneighbours.org/en/articles/china-unveils-n...

[2] https://english.www.gov.cn/news/202606/10/content_WS6a296017...

Re: Notes on DeepSeek

#49
post #8

"As a whole, China seems to treat AI as just another technology, rather than as some kind of singularity moment." This is a refreshing perspective.

The CCP knows, whatever the heck this technology will bring with itself, the current power dynamic inside of the country is on their side, and AI will solidify it. I hypothesize that, rather than slowly having it disperse in society and allow people to harness it in ways they don't want, they might as well accelerate everything until AI becomes the totalitarian swiss knife - which they can make use of in the best way…

US used AI (Claude on Maven) to determine a girl's elementary school as a target in war[0] and then triple tapped it and you're still more worried about hypothetical misuses of the single country responsible for this technology not being concentrated in the hands of a few powerful elite? ffs

[0] https://www.washingtonpost.com/national-security/2026/03/11/...

Re: Notes on DeepSeek

#50
post #41

Earlier quoted context omitted.

thanks. this really isnt that long, might as well paste in full here since OP deleted. Notes on DeepSeek: We visited the company HQ last Tuesday. It was founded in 2023 by Liang Wenfeng and operated out of his hedge fund, High-Flyer, until somewhat recently. The company released their R1 model in January 2025, so it was interesting to see what they’ve been doing The company is located in an unmarked, 12-story buildin…

I don't get the part of "AI models are not regulated directly, the government instead has restrictions on how those models can be used in software, services". Is it not the same thing? When I chat with DeepSeek about any (Chinese) political/social issue, it immediately begins aligning with the party's line or just cut off the conversation abruptly.

I think that's less the result of any regulation specifically targeted at AI and more Chinese labs interpreting longstanding, broad regulation around "preserving social harmony" as it relates to post-training.
Post reply on HN