[Addendum 2.23.26: After staring at this essay on and off over the weekend and playing with a few different prompts, I came up with a fourth requirement: the AI must be able to solve its own test without the answer key. This prevents trivial cryptographic-esque solutions that exploit the information asymmetry between the tester and the tested.]
Nice to see that the linked post brings this up. I was half tempted to nitpick that the test is trivially solved by asking us to invert a hash.
My results: no contradictions, three bullets bitten
My argument against Parfit’s depletion problem: the guilt of knowing that my living standards increase is at the cost of the well-being of future generations would cause my total welfare gain from such a policy to be around zero. If I didn’t care about non-existent people or was offered a ludicrously good life, I would bite the bullet and say it’s fine.
To clarify, I think it’s hard to say that it’s a moral wrong to harm future people, but it makes me feel bad so it’s a welfare loss.
(Also relevant SMBC on the repugnant conclusion)