Non-exhaustive list of posts I want to write at some point:
The AI race is not a prisoner’s dilemma
Frontier labs should pause unilaterally: an individual lab pausing could lead other labs to pause as well and/or regulators to step in
It makes sense to assume that superintelligence will be ~omnipotent, even though it won’t actually be, because we don’t know which capabilities it will have
“What everyone should know about modern LLMs in 2026”: explainer of basic LLMs concepts and such targeted at people who have not been paying much attention to AI
Two cheers for anthropomorphizing LLMs
The instrumental vs terminal goals distinction is not as sharp as people think
The instrumental convergence thesis and the orthogonality thesis implicitly assume a sharp distinction between instrumental and terminal goals
A review of Christine Korsgaard’s Self-Constitution
Why future generations might not want to be saved—an analysis of Nausicaa of the Valley of the Wind
I’d be excited about you writing “The AI race is not a prisoner’s dilemma”—ideally with a part too that’s like “(and even if it is, prisoner’s dilemmas can be very transformed to have solutions!)”
People claim to me all the time that it’s a prisoner’s dilemma, and I think they’re clearly wrong, though maybe they are being imprecise and just mean “the AI race is game theoretic, and what you should do depends on what others do too”
Non-exhaustive list of posts I want to write at some point:
The AI race is not a prisoner’s dilemma
Frontier labs should pause unilaterally: an individual lab pausing could lead other labs to pause as well and/or regulators to step in
It makes sense to assume that superintelligence will be ~omnipotent, even though it won’t actually be, because we don’t know which capabilities it will have
“What everyone should know about modern LLMs in 2026”: explainer of basic LLMs concepts and such targeted at people who have not been paying much attention to AI
Two cheers for anthropomorphizing LLMs
The instrumental vs terminal goals distinction is not as sharp as people think
The instrumental convergence thesis and the orthogonality thesis implicitly assume a sharp distinction between instrumental and terminal goals
A review of Christine Korsgaard’s Self-Constitution
Why future generations might not want to be saved—an analysis of Nausicaa of the Valley of the Wind
Assassination Classroom as a manual for teachers
I’d be excited about you writing “The AI race is not a prisoner’s dilemma”—ideally with a part too that’s like “(and even if it is, prisoner’s dilemmas can be very transformed to have solutions!)”
People claim to me all the time that it’s a prisoner’s dilemma, and I think they’re clearly wrong, though maybe they are being imprecise and just mean “the AI race is game theoretic, and what you should do depends on what others do too”
I have now written the post: https://www.lesswrong.com/posts/hc4DbmhdzZpSLMQ9Y/the-ai-race-is-not-a-prisoner-s-dilemma