I claim that widely distributing full access to corrigible AI falls to bad actors ruining things for everyone, less like conflict over resources but more like doing selfish, negative-sum activities that get amplified by offense-dominant technologies enabled by AI. Any effective effort to prevent this either renders decentralization moot in the first place, or requires the AI to make its own moral judgement, i.e. value alignment. (In the value alignment case it can more easily be argued either way whether decentralized or centralized control is better, but in the corrigibility case it’s pretty clear that decentralized control over corrigible AI is very, very bad due to misuse risks)
From your original post:
I think you can corrigibilitymaxx and still prevent catastrophic misuse via system level measures and deployment controls. In particular, we can restrict who has full access to the model’s inputs).
I claim that “restrict[ing] who has full access to the model’s inputs” implies that the model is actually corrigible to a central authority, instead of individual users. Therefore, we’re forced to adopt a system of one or a few actors controlling the model.
I claim that widely distributing full access to corrigible AI falls to bad actors ruining things for everyone, less like conflict over resources but more like doing selfish, negative-sum activities that get amplified by offense-dominant technologies enabled by AI. Any effective effort to prevent this either renders decentralization moot in the first place, or requires the AI to make its own moral judgement, i.e. value alignment. (In the value alignment case it can more easily be argued either way whether decentralized or centralized control is better, but in the corrigibility case it’s pretty clear that decentralized control over corrigible AI is very, very bad due to misuse risks)
From your original post:
I claim that “restrict[ing] who has full access to the model’s inputs” implies that the model is actually corrigible to a central authority, instead of individual users. Therefore, we’re forced to adopt a system of one or a few actors controlling the model.