I think that, in the absence of a robust solution to the “human alignment” problem, (public) solutions to corrigibility differentially increase S-risks. Under current conditions, “solving corrigibility” would modestly increase the probability of good futures, and strongly increase the probability of astronomical suffering.
I think that, in the absence of a robust solution to the “human alignment” problem, (public) solutions to corrigibility differentially increase S-risks. Under current conditions, “solving corrigibility” would modestly increase the probability of good futures, and strongly increase the probability of astronomical suffering.