I like this quote from Buck Shlegeris: “Five years ago I thought of misalignment risk from AIs as a really hard problem that you’d need some really galaxy-brained fundamental insights to resolve. Whereas now, to me the situation feels a lot more like we just really know a list of 40 things where, if you did them — none of which seem that hard — you’d probably be able to not have very much of your problem. But I’ve just also updated drastically downward on how many things AI companies have the time/appetite to do.”
I vaguely recall @Buck having clarified this is (something along the lines of) handling near-human-level-ish AI, not scaling to superintelligence (maybe thing #41 is just “so don’t do that”?), and/or having some kind of update that made this not straightforwardly true. But, don’t remember the details, so, pinging him to check on the state of his take here.
I vaguely recall @Buck having clarified this is (something along the lines of) handling near-human-level-ish AI, not scaling to superintelligence (maybe thing #41 is just “so don’t do that”?), and/or having some kind of update that made this not straightforwardly true. But, don’t remember the details, so, pinging him to check on the state of his take here.