Can you simply brainwash an LLM?
gradientdefense.com
Can you simply brainwash an LLM?
1–10 of 73 posts
Re: Can you simply brainwash an LLM?
#2Re: Can you simply brainwash an LLM?
#3Re: Can you simply brainwash an LLM?
#4Does this mean that I could train an LLM to do something like spread fake news? Would that even scale?
Re: Can you simply brainwash an LLM?
#5Re: Can you simply brainwash an LLM?
#6Does this mean that I could train an LLM to do something like spread fake news? Would that even scale?
Re: Can you simply brainwash an LLM?
#7Human-centric example but you get the point.
Re: Can you simply brainwash an LLM?
#8Does this mean that I could train an LLM to do something like spread fake news? Would that even scale?
Isn't this done with every "sanitized" LLM? Fake news is all according to perspective!
Re: Can you simply brainwash an LLM?
#9Does this mean that I could train an LLM to do something like spread fake news? Would that even scale?
Re: Can you simply brainwash an LLM?
#10While I’m sure they’re right - factually tampering with an LLM is possible - I doubt that this will be a widespread issue.
Using an LLM knowingly to generate false news seems like it will have similar reach to existing conspiracy theory sites. It doesn’t seem likely to me that simply having an LLM will make theorists more mainstream. And intentional use wouldn’t benefit from any amount of certification.
As far as unknowingly using a tampered LLM, I think it’s highly unlikely that someone would accidentally implement a model at meaningful scale which has factual inaccuracies. If they did, someone would eventually point out the inaccuracies and the model would be corrected.
My point is that an AI certification process is probably useless.