Came here to say this. It also seems worrying that the authors didn’t even mention this downside. Perhaps they’ll discuss it in the forthcoming blog posts but this is still a unilateralist’s curse situation.
I suspect that internally AI companies have benchmarks similar to this to hill climb onto, but yeah, yeesh, this is the exact opposite of safe.
Came here to say this. It also seems worrying that the authors didn’t even mention this downside. Perhaps they’ll discuss it in the forthcoming blog posts but this is still a unilateralist’s curse situation.
I suspect that internally AI companies have benchmarks similar to this to hill climb onto, but yeah, yeesh, this is the exact opposite of safe.