I have no idea. The conceit of AI alignment is that we will figure out how to answer that, somehow. I don’t expect us to solve alignment and I don’t even really know what it would mean to have solved it, but OP is written with the assumption that we do, and trying to reason about what that might imply.
I think that’s a really narrow view of what people mean by “alignment”. As far as I can tell, there’s no useful agreement, but you definitely don’t have to control something to have it implement values compatible with your own, or even exactly your own “real” values. If something is out there putting your CVE into practice, that doesn’t imply that you control it in any way.
The vagueness of the concept of “alignment” is a good reason not to use the word. I accept that there are a lot of people out there who think humans (however they define that) have to keep giving the orders forever, but I definitely don’t think that view can claim to own the term.
… and if it does, well… maybe you can have a superintelligence take your orders, but that doesn’t mean you’ll understand the details or the consequences, and I think that puts you on pretty shaky ground talking about “control”. Real control is fundamentally impossible and not something to waste time thinking about.
The what?
Regardless of how aligned it is, and whatever you mean by “aligned”, how do you control something you almost definitionally do not understand?
I have no idea. The conceit of AI alignment is that we will figure out how to answer that, somehow. I don’t expect us to solve alignment and I don’t even really know what it would mean to have solved it, but OP is written with the assumption that we do, and trying to reason about what that might imply.
I think that’s a really narrow view of what people mean by “alignment”. As far as I can tell, there’s no useful agreement, but you definitely don’t have to control something to have it implement values compatible with your own, or even exactly your own “real” values. If something is out there putting your CVE into practice, that doesn’t imply that you control it in any way.
The vagueness of the concept of “alignment” is a good reason not to use the word. I accept that there are a lot of people out there who think humans (however they define that) have to keep giving the orders forever, but I definitely don’t think that view can claim to own the term.
… and if it does, well… maybe you can have a superintelligence take your orders, but that doesn’t mean you’ll understand the details or the consequences, and I think that puts you on pretty shaky ground talking about “control”. Real control is fundamentally impossible and not something to waste time thinking about.