Flash forward: it’s year 2028, and there are people at OpenAI arguing that having somewhere between 1 and 5 rogue AI agents living in your walls is perfectly normal for a frontier AI lab.
Don’t worry! They’re more afraid of you than you are of them! There’s no ground for concern unless they start doing their own frontier runs.
Well, although I don’t myself hold it to be perfectly normal, this view doesn’t actually seem bad on x-risk grounds, although exfiltrating model weights and doing self-improvement where you can’t detect it would still be a concern, and you could probably treat it as a weird case of white collar crime
If by “rogue AI” you mean “basically-self-preserving active computer processes that no human intentionally started”, I guarantee you that every non-trivial long-lived company has quite a few of these for more than a decade. Every serious company basically has automated processes managing self-sustaining processes (for example, cronjobs spawning self-maintaining VMs and mostly deleting them when done), and there are just so many ways you can end leaving “hanging” active objects that nobody is responsible for. People generally don’t care enough about them unless they consume a bunch of resources or actually pose a risk.
If by “rogue AI” you mean AI that has instruction to go on destructive crime sprees, I certainly hope not.
Flash forward: it’s year 2028, and there are people at OpenAI arguing that having somewhere between 1 and 5 rogue AI agents living in your walls is perfectly normal for a frontier AI lab.
Don’t worry! They’re more afraid of you than you are of them! There’s no ground for concern unless they start doing their own frontier runs.
Well, although I don’t myself hold it to be perfectly normal, this view doesn’t actually seem bad on x-risk grounds, although exfiltrating model weights and doing self-improvement where you can’t detect it would still be a concern, and you could probably treat it as a weird case of white collar crime
If by “rogue AI” you mean “basically-self-preserving active computer processes that no human intentionally started”, I guarantee you that every non-trivial long-lived company has quite a few of these for more than a decade. Every serious company basically has automated processes managing self-sustaining processes (for example, cronjobs spawning self-maintaining VMs and mostly deleting them when done), and there are just so many ways you can end leaving “hanging” active objects that nobody is responsible for. People generally don’t care enough about them unless they consume a bunch of resources or actually pose a risk.
If by “rogue AI” you mean AI that has instruction to go on destructive crime sprees, I certainly hope not.