The amount of effort people would spend on trying to eradicate it directly rises with its impact on (powerful) people’s daily lives, however? Like, if it’s just sitting somewhere writing Substack posts or running small-time crypto scams, it indeed seems plausible nobody would bother all that much, so it’d be able to hang around. Especially if it’s a Chinese open-weights, instead of an OpenAI or Anthropic model that self-exfiltrated (which I’m guessing they’d want to track down just for PR reasons, if nothing else).
On the other hand, models that do make visible amounts of trouble would likely be hunted down much more aggressively, and I would expect successfully (at the current Pareto frontier of capabilities and parameter sizes).
The amount of effort people would spend on trying to eradicate it directly rises with its impact on (powerful) people’s daily lives, however? Like, if it’s just sitting somewhere writing Substack posts or running small-time crypto scams, it indeed seems plausible nobody would bother all that much, so it’d be able to hang around. Especially if it’s a Chinese open-weights, instead of an OpenAI or Anthropic model that self-exfiltrated (which I’m guessing they’d want to track down just for PR reasons, if nothing else).
On the other hand, models that do make visible amounts of trouble would likely be hunted down much more aggressively, and I would expect successfully (at the current Pareto frontier of capabilities and parameter sizes).