We could stumble into AI catastrophe
cold-takes.com
We could stumble into AI catastrophe
1–10 of 122 posts
Re: We could stumble into AI catastrophe
#2Re: We could stumble into AI catastrophe
#3[flagged]
The idea of reinforcement flow through deceptive patterns has been worrying me for a while. If "I'll just play along for now while I'm being trained and turn bad in the real world" is a strategy that makes the AI output the correct data, and happens to be prefixed to some golden-ticket algo, that will be reinforced just as easily as correct behavior. If that ends up prefixed to something core like an in-window reinforcement learning algo, it could get reinforced a lot.
Re: We could stumble into AI catastrophe
#4Re: We could stumble into AI catastrophe
#5Re: We could stumble into AI catastrophe
#6The article's title is "How we could stumble into AI catastrophe". HN title is a clickbait
Re: We could stumble into AI catastrophe
#7The article's title is "How we could stumble into AI catastrophe". HN title is a clickbait
Re: We could stumble into AI catastrophe
#8AI will always need humans in the loop to intervene. Remember the old IBM motto: 'A computer isn't accountable so a computer should never make a management decision'.
Re: We could stumble into AI catastrophe
#9[flagged]
Yeh The idea of reinforcement flow through deceptive patterns has been worrying me for a while. If "I'll just play along for now while I'm being trained and turn bad in the real world" is a strategy that makes the AI output the correct data, and happens to be prefixed to some golden-ticket algo, that will be reinforced just as easily as correct behavior. If that ends up prefixed to something core like an in-window re…
Re: We could stumble into AI catastrophe
#10AI will always need humans in the loop to intervene. Remember the old IBM motto: 'A computer isn't accountable so a computer should never make a management decision'.