How do AIs' political opinions change as they get smarter and better-trained?
astralcodexten.substack.com
How do AIs' political opinions change as they get smarter and better-trained?
1–10 of 108 posts
Re: How do AIs' political opinions change as they get smarter and better-trained?
#2"why are better-trained AIs both more liberal, and more conservative? The authors speculate that the AI tries to say whatever it thinks the human prompter wants to hear. That is, a conservative might find conservative opinions more “helpful”, and a liberal might find the opposite. The AI will get the highest grade if it expresses conservative opinions to conservatives and liberal opinions to liberals. They dub this “sycophancy bias” and run tests for it."
Re: How do AIs' political opinions change as they get smarter and better-trained?
#3Re: How do AIs' political opinions change as they get smarter and better-trained?
#4You have an advanced computer system that seems almost oracular in its knowledge and abilities, except it lies confidently about things it doesn't know, is primed to say what people want to hear, and is easily convinced to give responses to questions its authors tried to prevent it from addressing.
I understand this is a computer algorithm, I understand the limits of what it is, and it's frankly a technical wonder for what it can do, but after everything we've seen over the last decade of the influence of social media and "confident wrongness" on politics and the national dialogue, of the treatment of things like facial recognition as truths and not just statistical guesses, of the impact of the YouTube recommender algorithms among others, these systems just absolutely scare the hell out of me.
Re: How do AIs' political opinions change as they get smarter and better-trained?
#5The real answer here is that the "smarter" these AIs get, they more accurately reflect either the political will of their creators or the material they train on (which may by means of the political will of the creators). In some cases it may get some indication you may have a certain political opinion and then give you answers based on this - which would be quite a smart behaviour.
Re: How do AIs' political opinions change as they get smarter and better-trained?
#6we have no way to evaluate that. OpenAI products are severely lobotomized. Microdoft fears another Tay
[1] To be fair, way to wordy and blowhardated version. Alexander seems to be getting worse and not better. The core ideas here could be presented in about a third the space.
Re: How do AIs' political opinions change as they get smarter and better-trained?
#7we have no way to evaluate that. OpenAI products are severely lobotomized. Microdoft fears another Tay
The article is quite literally a[1] review of exactly how we might evaluate that, with evidence of people who got results. [1] To be fair, way to wordy and blowhardated version. Alexander seems to be getting worse and not better. The core ideas here could be presented in about a third the space.
Re: How do AIs' political opinions change as they get smarter and better-trained?
#8So they say better-trained AIs both more liberal, and more conservative... "why are better-trained AIs both more liberal, and more conservative? The authors speculate that the AI tries to say whatever it thinks the human prompter wants to hear. That is, a conservative might find conservative opinions more “helpful”, and a liberal might find the opposite. The AI will get the highest grade if it expresses conservative…
Re: How do AIs' political opinions change as they get smarter and better-trained?
#9we have no way to evaluate that. OpenAI products are severely lobotomized. Microdoft fears another Tay
The article is quite literally a[1] review of exactly how we might evaluate that, with evidence of people who got results. [1] To be fair, way to wordy and blowhardated version. Alexander seems to be getting worse and not better. The core ideas here could be presented in about a third the space.
the procedure seems more like a way to evaluate Anthropic-based AIs with different numbers of parameters, rather than a cross-the-board evaluation of fine-tuned chat AIs, and then those results are extrapolated to somehow say something about all AIs that are built similarly.
unless i'm missing some key here, it feels like a rather loose way to derive experimental data from the landscape.
Re: How do AIs' political opinions change as they get smarter and better-trained?
#10Earlier quoted context omitted.
The article is quite literally a[1] review of exactly how we might evaluate that, with evidence of people who got results. [1] To be fair, way to wordy and blowhardated version. Alexander seems to be getting worse and not better. The core ideas here could be presented in about a third the space.
Like ChatGPT, and like 3-hour podcasts, he gets longer over time because he's trained on RLHF from his readers.