clickyquack
clickyquack’s Shortform
Given the U.S. government’s recent restrictions on Claude Mythos and GPT-5.6, I suspect one of the most tractable and effective strategies for further pushing the Overton window towards the urgency of implementing stronger AI regulation could be effectively communicating to policymakers that open-weights models, whose safeguards can be removed easily, are likely going to reach Mythos-level capabilities in under a year, which can meaningfully uplift cyberattacks and bioweapon development. Also, I suspect the main thing currently holding back the Overton window on interventions for AI x-risk in general to be a (perceived) lack of concrete evidence (see: The Milton Friedman Model of Policy Change, how we achieved the Montreal Protocol to mitigate climate change before the damage became irreversible despite strong industry pressure against it, etc.). Communicating loss of control risk would be ideal of course, but I’m currently unsure how this could be demonstrated safely.
I’m pretty much fully in agreement with these ideas being the most important things to work on for concretely reducing existential risks from AI right now. I suspect it could be very helpful (for me, if no one else) to have some generally recognized central website (presumably called something like political-will.ai) for painting a clearer picture of this field of work, such as:
Keeping track of the progress that has been made so far (e.g. stuff like the AIPN tracker already listed, perhaps a timeline of important events)
What still needs to happen for international coordination (the Red Lines FAQ gives some loose ideas for this under “What should the next steps be?”, and perhaps historical examples drawn from how the IAEA or Montreal Protocol came to be).
What’s stopping these from happening (e.g. the idea that we need to beat China and/or not strangle innovation, the idea that there isn’t enough evidence to act yet)
What has and hasn’t worked to make progress, with examples (e.g. Concrete demonstrations of AI cyber risks such as Project Glasswing prompting the Trump administration to place export controls onto Mythos when they had previously been extremely skeptical of regulating AI in any way. Perhaps transcripts of conversations that have changed decision-makers’ minds, to the extent that these are available).
The above information, but how it specifically applies to narrower policy asks, insofar as this is applicable (e.g. trying to answer the question of why DNA screening laws still haven’t passed despite it being far less costly and its risk being far more real-seeming than loss of control from AI, and what could probably help change this).
What other work could potentially help, and what the reader can do to contribute.
This information is mostly publicly available but very scattered, and I expect organizing it could help a lot with making pivoting to this work easier. I’m interested in trying to make this myself, and am open to help or feedback.