Also known as Max Harms. (I post AI alignment content under my other account.)
Not the same person as MaxH!
Raelifin
Those Who Make History
Planning for Preservation in the Age of AI
Max’s stopgap plan would involve comparing the human’s values to what they’d be if the AI did nothing. I think he understates how bad that stopgap plan is. Even providing straightforwardly-true factual information can change what a person wants, right?
Yes. That’s right. And I am (among other things) worried about an AI that warps my values by telling me a series of facts.
But I want to clarify that I’m talking about terminal values, not strategic sub-goals. The stopgap plan is 100% able to tell me that the store is closed, thus changing my plan of going to the store. What it shouldn’t do is tell me intense stories about the suffering of pigs and thereby change how much I care about pigs.[1]
Why do you think this stopgap is so bad? (I agree that it’s bad, but it seems like you see it as worse than me.)
- ^
Unless this story is necessary to counteract another pressure such that I cleave closer to the null-action counterfactual.
- ^
The dynamic you’re talking about is real, but also I suspect is a product of the Overton window being closed. Marketing and building momentum could, I think, unlock a huge market. But absent doing better than the existing orgs on that front, I agree the customer base is likely going to be tiny.
I don’t have a good sense of how well MAiD cryonics can get in terms of information preservation. On the surface it should be a huge improvement, since it won’t have the ischemia issue, but three immediate things jump to mind:
* Ken Hayworth effectively condemned the highest-quality cryo-tissue in 2015, not just the average case (which he agrees is much worse).
* Cryo orgs store at −196, and there is a substantial risk of shattering at that temperature.
* Being unable to survive thawing means cryo is more fragile in many ways.
Thanks for the response! I appreciate the extra context on your third-party validation. Sorry if that’s on the website somewhere and I just didn’t find it.
I actually learned about Sparks in the course of investigating Nectome, and I hope my post helps similarly direct more attention your way. I broadly think Sparks is doing good work in saving lives, and I appreciate your contribution to that. :)
If you succeed, maybe you should work for Nectome. 😅
Mostly our conversation was about MAiD, and the way that donating your body as part of that can require crossing country borders. But honestly I don’t remember a ton of specifics because I wasn’t taking notes during that exchange. Maybe @Borys Wrobel has more to say. (Tho he doesn’t use LW that much, so if you really want to know, you might want to email Nectome. Hello@nectome.com)
Yeah, totally. I expect historical records and the memories of other people to be useful.
My point is that I don’t know an objective measure for whether the superintelligence rescued the existing person or built a new person, except via whether they match the other memories and records. If the superintelligence optimizes for the “rescued” person matching the memories of those who knew them, they will seem like they were revived successfully, but might not actually be very close to the real deal.
Yes.
Also @Aurelia is on LW and might be willing to answer questions herself.
Hard to say what the future can/can’t do. I think I’m, like, 80% that a brain that’s simply dumped in LN2 is going to lose so much information that even a superintelligence could not put the person back together in a way that their loved ones think they simply came back from the dead without being fundamentally changed (modulo cheating by “repairing” the cryonaut in a way that is deliberately designed to match the memories of their loved ones). Like, at the far end of what might be the case, the frozen brain tissue might as well have gone into the fire. The superintelligence can build a person that matches the historical record, but they won’t be the same person.
It could also be the case that the relevant information is still there, even when shredded, like papers put through a shredder, and that a sufficiently dedicated agent could figure out a model of how the ice formed, simulate an inverse process, and have things be fine. Even if I get in an accident and I’m at room temp for days, I would still like to be cryopreserved just in case this is true. But I wouldn’t bet on it.
In the common cryo case, it gets even trickier, since some parts of the brain will be well-perfused, and others won’t be, and there’s a quantitative question of how much. If I lost 10% of my cortex I would still be pretty similar, but would also be pretty different. I don’t think we have good measures here.
In short: idk, my guess is that reality is complicated and “invert the shredding” is not as simple as it sounds, even if it’s possible, in some sense.
Nectome: All That I Know
Raelifin’s Shortform
Does anyone have any questions that they’d like me to ask Nectome? I’m visiting their facilities on Wednesday and getting some VIP access. I think they’re quite happy to answer questions directly, but since I’m doing a deep-dive on them, there may be things that I, as an outsider, am more capable of answering as part of my investigation.
The only thing I can think of is Three Worlds Collide, but that’s by Eliezer, and doesn’t exactly fit your description. https://www.lesswrong.com/posts/HawFh7RvDM4RyoJ2d/three-worlds-collide-0-8
This is great. I don’t remember the last time a post made me flip so hard from “the thesis is obviously false” to “the thesis is obviously true”! And not just because I didn’t understand it, but also because I learned a thing. (Tho part was a pedagogically useful misunderstanding.)
Nope. If you have specific questions I’d be happy to answer them.
Becoming a Chinese Room
This market looks semi-reasonable to me: https://manifold.markets/MaxHarms/when-will-the-red-heart-audiobook-c
It’s hard for me to make concrete predictions because I have a lot of agency, and it depends on things like my standards and priorities. It turns out I’m moving across town in December, so that will delay things. If I had to guess, I would say the market is over-estimating the chance it’ll be <January and >April and under-estimating my chances of getting it done in February, but :shrug:.
Yeah. I think this is a CDT fallacy. We can influence the chance that the world ends quite a lot. “We” is made up of a bunch of “me”. Therefore, yes, I do think I can make a difference (as part of my decision group).
I would suggest asking if they vote, but a lot of people vote for non-FDT reasons (eg high expected impact even if you only have a 1/million chance of changing things).