Earlier quoted context omitted.
Sure, but humans are a relatively known entity. They exist with a level of variance that is not too extreme, and you can more or less, in a general way, predict their behavior to a tolerable risk level. We've seen billions of humans, we know what to expect of them. We might not understand how other humans work internally, but we understand their actionspace. For AI models, that isn't really the case. We don't know ho…
Then train the models on real world data? Verify outputs enough until confidence is achieved. The computer can do whatever it wants. People can do whatever they want. The question will be what level of security access will they have. The key difference today is people are really good at making rationalizations for individual decisions. Computers are not. Sometimes decisions are generally important, when they are impo…
And therein lies the rub. Do people actually want to know the truth? They invent methods to obtain it, but is that their aim or is it to confirm their current understanding of the world? Unlike a human being that can be forced into silence through coercion or manipulation, a computational model, once proven with certainty, is never going to go back in Pandora's box. This contradiction of pursuing truth but suppressing the inconvenience of its conclusions hasn't disappeared for some reason despite greater and greater knowledge of ourselves. Will it ever?