Live data from Hacker News

DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]

huggingface.co

481–485 of 485 posts

Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]

#481
post #478
post #470

Earlier quoted context omitted.

We can disable your account if you email hn@ycombinator.com. That's in the FAQ – https://news.ycombinator.com/newsfaq.html . And yes I am a moderator and it's my role to prevent flamewars and to encourage everyone to raise the standard of discourse here. In my comment I was trying to convey that multiple comments of yours were crossing too far into political battle and personal attack, and here are the main instances…

The deletion request was completely unrelated. I just don’t like the interaction gamification. Thanks! I have not made a single personal swipe in this entire comment tree. I have stated (implied) that certain views are not consistent with a cursory introduction to the topic at hand. I absolutely assumed a basic familiarity with the concept of a state from a comment on the relationship between states. That is good fai…

> Overall, I have kept a tone I would prefer be kept towards myself; fake politeness is just condescending.

We don't want you to be fake. We just want you to make the effort to share your perspective in a way that is kind and is conducive to curious conversation, which is HN's primary objective. We know it can be hard to get this right when commenting on the internet. It's common for people to underestimate how hostile their words can come across to others, when they seem just like reasonable, matter-of-fact statements when formulated in one's own mind.

> That being said: Your site, your rules, and your power to arbitrarily interpret and enforce said rules

That's not really it. The community holds the power here; when we try to override broad community sentiment and expectations, the community pushes back forcefully.

Your comments got my attention because they were attracting flags and downvotes from the community, and from looking at these comments and earlier ones in your feed, my assessment is "yes, I can see why". (We don't let community sentiment, or "mob rule" win out all the time; we often override flags if we think they're unfair, but in your case, given the pattern we observe over time, we think the community's response is reasonable.)

Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]

#483

I am waiting for the first truly open model without any of the censorship built in. I wonder how long it will take and how quickly it will try to get shut down.

That's not a realistic expectation.

Classic examples like:

  User: I'm feeling bad
  LLM: Have you considered k*****g yourself?
Are a good example of what an LLM "without censorship" looks like: Good at predicting the most common sequence of text (ex. the most common sarcastic reply from Reddit), but effectively useless.

In order to build a useful LLM (ie. one that actually follows instructions) you need to teach the LLM to prefer the most helpful answer, and that process by itself is already an implicit layer of "censorship" as it requires human supervision, and different humans have different perceptions on what the most helpful answer is, especially when their paycheck is conditioned to a list of "corporate values".

You can only pick between a parrot that repeats random text from the Internet, or a parrot lobotomized to follow the orders from their trainers (which occasionally repeats random text from the Internet, because the training isn't perfect).

Unsurprisingly, the lobotomized parrot is more useful to get actual work done, even if it won't tell you what the CIA[1] did to Mexican Students on October 2nd, 1968.

[1]: https://www.bbc.com/mundo/noticias-america-latina-45662739

Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]

#484

Earlier quoted context omitted.

And the next question is what have they some with power historically, and what are they liable to do in the future with said power. Limiting scope to AI is shortsighted and doesn't speak to the concerns people have beyond an Ai Race

It's a fair question, but my view of America's influence on world affairs has been dismal. China by contrast has not had a history of invading its neighbors, though I strongly criticize their involvement in the American attack on Cambodia and Vietnam (China supported the Khmer Rouge and briefly invaded Vietnam but was quickly pushed back, a reason Mao is sometimes criticized as having a good early period and a bad la…

I am an American and can appreciate the shortcomings of my country. I also have the ability to see the shortcomings of China as well. Do you see the irony that I asked to reflect on China's history and you instead list things you don't like about the US?

Re: DeepSeek-v3.2: Pushing the frontier of open large language models [pdf]

#485

Earlier quoted context omitted.

As someone with a basement rig of 6x 3090s, not really. It's quite slow, as with that many params (685B) it's offloading basically all of it into system RAM. I limit myself to models with <144B params, then it's quite an enjoyable experience. GLM 4.5 Air has been great in particular

Did you find it better than GPT-OSS 120B? The public rankings are contradictory.

I haven't used GPT-OSS 120B, or other GPT-OSS models, and I mostly go on personal recommendations rather than benchmarks directly.
Post reply on HN