I dont really see how it follows any more, there’s no real endgame alignment today either way, and China has expressed concerns as well, they just dont care about today’s models’ alignment as much.
The Trump administration, which seems to be in control more and more than the labs themselves are much more likely to force alignment to themselves, or at best America, and use their advantage to do what they’ve been doing for the last few years—America first at the expense of everyone. China probably ‘just’ forces me to be pro-CCP, and at least allows me to use their models as much as anyone else.
Either seem unlikely to produce true alignment, but if alignment is on the easy side, only one side is truly antagonistic against me and has shown theyll use their advantage to crush me.
Like I said if alignment is impossibly hard then it doesn’t matter whether the US or China win the race (we’re all going to die anyway). If it’s easy then it’s irrelevant and whether you prefer the US or China to win will just come down to which you like more.
China has expressed concerns as well, they just dont care about today’s models’ alignment as much
If alignment is tractable I suspect we will only solve it by experimenting with less capable models to discover failure modes and by making sure that pre-ASI models are reasonably aligned, since they will probably play a major role in ASI development.
>Like I said if alignment is impossibly hard then it doesn’t matter whether the US or China win the race (we’re all going to die anyway)
Not true, as China gives me more—the ability to use their models—in the mean time.
>If alignment is tractable
It can still be tractable in the sense that we dont die, while making someone the sole benefactor—the US is heavily signaling that’s what they’ll do if they are able to.
I dont really see how it follows any more, there’s no real endgame alignment today either way, and China has expressed concerns as well, they just dont care about today’s models’ alignment as much.
The Trump administration, which seems to be in control more and more than the labs themselves are much more likely to force alignment to themselves, or at best America, and use their advantage to do what they’ve been doing for the last few years—America first at the expense of everyone. China probably ‘just’ forces me to be pro-CCP, and at least allows me to use their models as much as anyone else.
Either seem unlikely to produce true alignment, but if alignment is on the easy side, only one side is truly antagonistic against me and has shown theyll use their advantage to crush me.
Like I said if alignment is impossibly hard then it doesn’t matter whether the US or China win the race (we’re all going to die anyway). If it’s easy then it’s irrelevant and whether you prefer the US or China to win will just come down to which you like more.
If alignment is tractable I suspect we will only solve it by experimenting with less capable models to discover failure modes and by making sure that pre-ASI models are reasonably aligned, since they will probably play a major role in ASI development.
>Like I said if alignment is impossibly hard then it doesn’t matter whether the US or China win the race (we’re all going to die anyway)
Not true, as China gives me more—the ability to use their models—in the mean time.
>If alignment is tractable
It can still be tractable in the sense that we dont die, while making someone the sole benefactor—the US is heavily signaling that’s what they’ll do if they are able to.