Live data from Hacker News

Machine Unlearning in 2024

ai.stanford.edu

1–10 of 97 posts

Re: Machine Unlearning in 2024

#3
> However, RTBF wasn’t really proposed with machine learning in mind. In 2014, policymakers wouldn’t have predicted that deep learning will be a giant hodgepodge of data & compute

Eh? Weren't deep learning and big data already things in 2014? Pretty sure everyone understood ML models would have a tough time and they still wanted RTBF.

Re: Machine Unlearning in 2024

#4
Why should we try to unlearn "bad" behaviours from AI?

There is no AGI without violence, its part of being free thinking and self survival.

But also by knowing that launching a first strike by a drunk president was a bad idea we averted a war because of a few people, AI needs to understand consequences.

It seems futile to try and hide "bad" from AI.

Re: Machine Unlearning in 2024

#5
post #3

> However, RTBF wasn’t really proposed with machine learning in mind. In 2014, policymakers wouldn’t have predicted that deep learning will be a giant hodgepodge of data & compute Eh? Weren't deep learning and big data already things in 2014? Pretty sure everyone understood ML models would have a tough time and they still wanted RTBF.

I don't know if people anticipated contemporary parroting behavior over huge datasets. Modern well funded models can recall an obscure persons home address buried deep into the training set. I guess the techniques described might be presented to the European audience in an attempt to maintain access to their data/and or market for sales. I hope they fail.

Re: Machine Unlearning in 2024

#6

Why should we try to unlearn "bad" behaviours from AI? There is no AGI without violence, its part of being free thinking and self survival. But also by knowing that launching a first strike by a drunk president was a bad idea we averted a war because of a few people, AI needs to understand consequences. It seems futile to try and hide "bad" from AI.

This is presumably about a chatbot though, not AGI, so it's basically a way of limiting what they say. (Not a way that I expect to succeed)

Re: Machine Unlearning in 2024

#7
“to edit away undesired things like private data, stale knowledge, copyrighted materials, toxic/unsafe content, dangerous capabilities, and misinformation, without retraining models from scratch”

To say nothing of unlearning those safeguards and/or “safeguards”.

Re: Machine Unlearning in 2024

#8

Why should we try to unlearn "bad" behaviours from AI? There is no AGI without violence, its part of being free thinking and self survival. But also by knowing that launching a first strike by a drunk president was a bad idea we averted a war because of a few people, AI needs to understand consequences. It seems futile to try and hide "bad" from AI.

So you have a problem with supervised learning like spam classifiers?

Re: Machine Unlearning in 2024

#9

Why should we try to unlearn "bad" behaviours from AI? There is no AGI without violence, its part of being free thinking and self survival. But also by knowing that launching a first strike by a drunk president was a bad idea we averted a war because of a few people, AI needs to understand consequences. It seems futile to try and hide "bad" from AI.

Because we can get AI related technologies to do things living creatures can’t, like provably forget things. And when it benefits us, we should.

Personal opinion, but I think AGI is a good heuristic to build against but in the end we’ll pivot away. Sort of like how birds were a good heuristic for human flight, but modern planes don’t flap their wings and greatly exceed bird capabilities in many ways.

Attribution for every prediction and deletion seem like prime examples of things which would break the analogy of AI/AGI with something more economically and politically compelling/competitive.

Re: Machine Unlearning in 2024

#10
I've wondered before if it was possible to unlearn facts, but retain the general "reasoning" capability that came from being trained on the facts, then dimensionality reduce the model.
Post reply on HN