How will the Google/Anthropic/OpenAI's of the world make money on AI if open models are competitive with their models? What hurt open source in the past was its inability to keep up with the quality and feature depth of closed source competitors, but models seem to be reaching a performance plateau; the top open weight models are generally indistinguishable from the top private models. Infrastructure owners with acce…
I call this the "Karl Marx Fallacy." It assumes a static basket of human wants and needs over time, leading to the conclusion competition will inevitably erode all profit and lead to market collapse. It ignores the reality of humans having memetic emotions, habits, affinities, differentiated use cases & social signaling needs, and the desire to always want to do more...constantly adding more layers of abstraction in…
DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]
271–280 of 485 posts
Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]
#272Earlier quoted context omitted.
[flagged]
> For example, a small random percentage of the time, it could add a subtle security vulnerability to any code generation. Now on the HN frontpage: "Google Antigravity just wiped my hard drive" Sure going to be hard to distinguish these Chinese models' "intentionally malicious actions"! And the cherry on top: - Written from my iPhone 16 Pro Max (Made in China)
Even if China did manage to embed software on the iPhone in Taiwan, it would soon hopefully be wiped since you usually end up updating the OS anyway as soon as you activate it.
Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]
#273Earlier quoted context omitted.
Competitor != adversary. It is US warmongering ideology that tries to equate these concepts.
[flagged]
The topic itself, like any topic, is fine to discuss here, but care must be taken to discuss it in a de-escalatory way. The words you use and the way you use them matter.
Most importantly, it's not OK to write "it is however entirely reasonable to assume that the comment I replied to was made entirely in bad faith". That's a swipe and a personal attack that, as the guidelines ask, should be edited out.
Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]
#274Earlier quoted context omitted.
Big enterprise with mostly private companies as their clients? Lol, yeah, that’s how they work from my personal experience. The reality is, if it’s not a tech-first enterprise and already outsource part of tech to a shop outside of NA (which is almost majority at this point), they will do absolutely everything to cut the costs.
I spent three years working in consulting mostly in public sector and education and the last two working with startups to mid size commercial interest and a couple of financial institutions. Before that I spent 6 years working between 3 companies in health care in a tech lead role. I’m 100% sure that any of those companies would I have immediately questioned my judgment for suggesting DeepSeek if had been a thing. Ab…
Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]
#275Earlier quoted context omitted.
[flagged]
Competitor != adversary. It is US warmongering ideology that tries to equate these concepts.
Please don't engage in political battle here, including singling out a country for this kind of criticism. No matter how right you are or feel you are, it inevitably leads to geopolitical flamewar, which has happened here.
Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]
#276Earlier quoted context omitted.
Competitor != adversary. It is US warmongering ideology that tries to equate these concepts.
you clearly haven't been paying attention remember when the US bugged EU leader's phones, including Merkel from 2002 to 2013?
Please don't be snarky or condescending in HN comments. From the guidelines: Be kind. Don't be snarky. Converse curiously; don't cross-examine. Edit out swipes.
Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]
#277Earlier quoted context omitted.
No… Nobody I work for will touch these models. The fear is real that they have been poisoned or have some underlying bomb. Plus y’know, they’re produced by China, so they would never make it past a review board in most mega enterprises IME.
People say that, but everyone, including enterprises, are constantly buying Chinese tech one way or another because of cost/quality ratio. There’s a tipping point in any excel file where risks don’t make sense, if the cost is 20x for the same quality. Of course you’ll always have exceptions (government, military and etc.), but for private, winner will take it all.
Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]
#278I hate that their model ids don't change as they change the underlying model. I'm not sure how you can build on that. % curl https://api.deepseek.com/models \ -H "Authorization: Bearer ${DEEPSEEK_API_KEY}" {"object":"list","data":[{"id":"deepseek-chat","object":"model","owned_by":"deepseek"},{"id":"deepseek-reasoner","object":"model","owned_by":"deepseek"}]}
Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]
#279Why are there so few 32,64,128,256,512 GB models which could run on current consumer hardware? And why is the maximum RAM on Mac studio M4 128 GB??
Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]
#280Earlier quoted context omitted.
I spent three years working in consulting mostly in public sector and education and the last two working with startups to mid size commercial interest and a couple of financial institutions. Before that I spent 6 years working between 3 companies in health care in a tech lead role. I’m 100% sure that any of those companies would I have immediately questioned my judgment for suggesting DeepSeek if had been a thing. Ab…
I've worked with financial services, and insurance providers that would have done the opposite for cost saving measures. So, I'm not sure what to say here.