Earlier quoted context omitted.
>Just because humans make mistakes (or even as many or more mistakes than machines) doesn't mean the nature or consequences of the mistakes are the same. True, but this shouldn't be hard to measure. It's not like we will go from these AIs will go from teaching no one to replacing all teachers overnight. Give a random sample of classes access to a GPT-n powered tutor, and see if they do better or worse over the next f…
> this shouldn't be hard to measure > see if they do better or worse over the next few years This doesn't strike me as entirely ethical. Also, it may be hard to gauge the long term effects of subtle misinformation generated by these models.
That said, you have to compare results from different strategies, or else you'll never improve or react to changing circumstances. I think it's like medical trials; you experiment, but be willing to stop the moment you get good evidence that one option is worse than another.