I’m interested in doing in-depth dialogues to find cruxes. Message me if you are interested in doing this.
I do alignment research, mostly stuff that is vaguely agent foundations. Currently doing independent alignment research on ontology identification. In 2023 I was on Vivek’s team at MIRI, before that I did MATS 2, and before that I did a CS and Neuroscience undergrad (thesis on statistical learning theory).
The best summary of my AI & alignment beliefs is my corrigibility basin of attraction post from 7 months ago.
Sorry to add to a pile-on, but just for the record:
I think the positive case for you joining OpenAI is weak-to-non-existent. Having a marginally better model of RSI isn’t clearly valuable. More detail and higher confidence parameters aren’t very important for governance or other interventions.
The case for negative impact is strong. You’ve said you know that people inside OpenAI want to use your work to accelerate RSI.
My primary model of your behavior is that you’re optimizing for importance and people paying attention to you (and maybe secondarily money), rather than primarily trying to do good. To the extent you consider me a friend, I hereby pressure you to quit.
I appreciate you writing this list of takes, I’ve heard most of them but it’s good to put things online and some of them are interesting. I disagree with some (particularly 7 and 8) for standard reasons we’ve discussed.