wedrifid comments on The AI in a box boxes you

wedrifid 4 Feb 2014 8:44 UTC
2 points
0

We can talk about precommitment all day, but the fact of the matter is that humans can’t actually precommit.

Pre-commitment isn’t even necessary. Note that the original explanation didn’t include any mention of it. Later replies only used the term for the sake of crossing an inferential gap (ie. allowing you to keep up). However, if you are going to make a big issue of the viability of precommitment itself you need to first understand that the comment you are replying to isn’t one.

That wasn’t a Causal Decision Theorist attempting to persuade someone that it has altered itself internally or via an external structure such that it is “precommited” to doing something irrational. It is a Timeless Decision Theorist saying what happens to be rational regardless of any previous ‘commitments’.

ur cognitive architectures don’t have that function. Sure, we can do our very best to act as though we can, but under sufficient pressure there are very few of us whose resolve will not break.

I’m aware of the vulnerability of human brains, so is Eliezer. In fact the vulnerability of human gatekeepers to influence even by humans, much less super-intelligences is something Eliezer made huge deal about demonstrating. However this particular threat isn’t a vulnerability of Eliezer or myself or any of the others who made similar observations. If you have any doubt that we would destroy the AI you have a poor model of reality.

It’s easy to convince yourself of having made an inviolable precommitment when you’re not actually facing e.g. torture.

For practical purposes I assume that I can be modified by torture such that I’ll do or say just about anything. I do not expect the tortured me to behave the way the current me would decide and so my current decisions take that into account (or would, if it came to it). However this scenario doesn’t involve me being tortured. It involves something about an AI simulating torture of some folks. That decision is easy and doesn’t cripple my decision making capability.