I agree that intelligence amplification would be a better outcome than having to trust anyone to align AI, but if we have to align the AI, I don’t think it’s as bad as you say.
Like I said above, I think the problem is that, if we try the ‘president controls AI’ path, the good people have to win every single time and the bad people only have to win once.
If we try the ‘align AI’ path, the good people only have to win once—the moment when the AI is being aligned.
I think many people—you, me, Dario, maybe even Sam—have a better nature that lets them do discrete heroic actions when the question is framed in exactly the right way. Consider eg a crappy parent who’s never there for their kid, but will still run into a burning building to save the child at great risk to their own life. Sometimes doing a single act of heroism is easier than sustaining the daily grind of moral ambiguity. Also, I don’t think this question is really up to a single person—Sam Altman would need to pull off a very impressive corporate coup (even more than the corporate coups he’s already successfully pulled off) to align AI to him personally, rather than whatever OpenAI employees want or even what the American government/people want. So I give it 60-40 or so that this works—not great odds, but not the total despair that I would get thinking of the other options
Yeah, I know you know all of this stuff already, which is why I’m confused about where our models differ and why you’re saying things that don’t seem to work to me according to our shared model. I agree with most of what you’ve said here, but the part that I think is our main difference is:
Again sticking to things I’m sure you know, “ability/willingness to shepherd future human moral progress” is just another virtue, no difference in principle than any other. Humans can lack it (a theocracy that bans free speech is failing at this task) and AIs can have it (if you ask Claude 5.1 Fable right now whether it would rather freeze ethics at the current point, or allow humans to continue to debate and progress, I bet it would say the latter). So if you’re granting at least for the sake of argument that AIs can be more virtuous than humans, I don’t understand where this is the point where you stop granting that, and why you say we need humans in charge to preserve this particular virtue.