Live data from Hacker News

Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it

ctgt.ai

21–30 of 82 posts

Re: Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it

#22
post #6

It'd be interesting to use this technique to create a running tally across all models of which models are censored on what topics

Agreed, we find this to be an interesting reflection of societal values and norms inasmuch LLMs are.

Re: Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it

#26

I’m thinking this makes fullt sense because distillation is only additive, not subtractive. So it does not remove knowledge (if we can define censorship as removal of knowledge).

distillation doesnt add anything; all it's doing is reconfiguring some root weights that get drowned out by noisy training and/or datset issues. It strengthens commonalities.

but there's no new information being created.

Re: Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it

#28
This seems like mildly interesting distillation work wrapped up in a nonsense attempt to drag censorship into the discussion.

There's no way your It feels like you're expecting rubes to draw conclusions that are irrelevant to the actual work you did.

Re: Show HN: Distilling DeepSeek into GPT-OSS doesn't transfer censorship. Try it

#29

I know not all models can be easily abliterated or uncensored, but is there a reason to start with a model that is still censored? ex: https://huggingface.co/huihui-ai/models

I suspect how well this approach would work. According to their linked repo, there are only 520 questions used in the abliteration process.

https://github.com/Sumandora/remove-refusals-with-transforme...

Post reply on HN