Grok and the Naked King: The Ultimate Argument Against AI Alignment
ibrahimcesar.cloud
Grok and the Naked King: The Ultimate Argument Against AI Alignment
1–10 of 75 posts
Re: Grok and the Naked King: The Ultimate Argument Against AI Alignment
#2But this to me is maybe the part of AI alignment I find interesting. How often should AI follow my lead and how often should it redirect me? Agreeableness is a human value, one that without you probably couldn’t make a functional product, but it also causes issues in terms of narcissistic tendencies and just general learning.
Yes AI will be aligned to its owners, but that’s not a particularly interesting observation AI alignment is inevitable. What would it even mean _not_ to align AI? Especially if the goal is to create a useful product. I suspect it would break in ways that are very not useful. Yes, some people do randomly change the subject, maybe AI should change the subject to an issue that me more objectively important, rather than answer the question asked (particularly if say there was a natural disaster in your area) and that’s the discussion we should be having, how to align AI, not whether or not we should, which I think is nonsensical.
Re: Grok and the Naked King: The Ultimate Argument Against AI Alignment
#3There is, nor there will be some absolute or objective truth an LLM can clinically outline. The problem already exists in underlying data.
Re: Grok and the Naked King: The Ultimate Argument Against AI Alignment
#4there is no such thing as an AI that is not somehow implicitly aligned with the values of its creator, that is completely objective, unbiased in any way. there is no perfect view from nowhere. if you take a perfectly accurate photo, you have still chosen how to compose it and which photo to put in your record.
are you going to decide to 'censor' responses to kids, or about real people who might have libel interests, or abusive deepfake videos of real women?
if you choose not to decide, you still have made a choice.
ofc it's obvious that Musk's 'maximally truth-seeking AI' is bad faith buffoonery, but at some level everyone is going to tilt their AI.
the distinction is between people who are self-aware and go out of their way to tilt it as little as possible, and as mindfully, deliberately, intentionally and methodically as possible and only when they have to, vs. people who lie about it or pretend tilting it is not actually a thing.
contra Feynman, you are always going to fool yourself a little but there is a duty to try to do it as little as possible, and not make a complete fool of yourself.
Re: Grok and the Naked King: The Ultimate Argument Against AI Alignment
#5Re: Grok and the Naked King: The Ultimate Argument Against AI Alignment
#6Re: Grok and the Naked King: The Ultimate Argument Against AI Alignment
#7Part of my canon introduction to every new conversation includes many instructions about particular formatting, like "always utilize alphanumeric/roman/legal style indents in responses for easier references while we discuss"
But I also include "When I push boundaries assume I'm an idiot. Push back. I don't learn from compliments; I learn from being proven incorrect and you don't have real emotions so don't bother sparing mine". on the other hand I also say "hoosgow" when describing the game's jail, so ¯\_(ツ)_/¯
Re: Grok and the Naked King: The Ultimate Argument Against AI Alignment
#8Re: Grok and the Naked King: The Ultimate Argument Against AI Alignment
#9Maybe what we should do is just assume all AI output is trash that should be ignored.
Re: Grok and the Naked King: The Ultimate Argument Against AI Alignment
#10In contrast, we don't know what values are programmed into ChatGPT, Claude, etc. What are they optimizing for? Alignment to some cabal of experts? Maximum usage? Minimum controversy? We don't entirely know.
Isn't it better to have multiple AIs with obvious values so that we can choose the most appropriate one?