jacob_cannell comments on A caveat to the Orthogonality Thesis

jacob_cannell 10 Nov 2022 16:18 UTC
3 points
0
Humans have sophisticated world models that contain simulacra of other humans—specifically those we are most familiar with. I think you could make an interesting analogy perhaps to multiple personality disorder being an example of simulacra breaking out of the matrix and taking over the simulation, but it’s a bit of a stretch.

What are sharp gradients? Is that a well established phenomena in ML you could point me at?

The simulacra in a simulation do not automatically break out and takeover, that’s just a scary fantasy. Like so much of AI risk discourse—just because something is possible in principle does not make it plausible or likely in reality.
- Thane Ruthenis 10 Nov 2022 16:33 UTC
  2 points
  0
  Parent
  I’m not talking about a simulacrum breaking out.
  Uh, apologies, I meant steepest gradients. The SGD is a locally greedy optimization process that updates a ML model’s parameters in the direction of the highest local increase in performance, i. e. along the steepest gradients. I’m saying that once there’s a general-purpose problem-solving algorithm represented somewhere in the ML model, the SGD (or evolution, or whatever greedy selection algorithm we’re using) would by default attempt to loop it into the model’s own problem-solving, because that would increase its performance the most. Even if it was assembled accidentally, or as part of the world-model and not the ML model’s own policy.