Most of my posts and comments are about AI and alignment. Posts I’m most proud of, which also provide a good introduction to my worldview:
Without a trajectory change, the development of AGI is likely to go badly
Steering systems, and a follow up on corrigibility.
I also created Forum Karma, and wrote a longer self-introduction here.
PMs and private feedback are always welcome.
NOTE: I am not Max Harms, author of Crystal Society. I’d prefer for now that my LW postings not be attached to my full name when people Google me for other reasons, but you can PM me here or on Discord (m4xed) if you want to know who I am.
No, not really. I think moderation is unlike other areas you’ve been prescient, for a few reasons:
We already have examples of what the end state of lots more author moderation looks like. Namely Twitter / X, where there is a norm of individuals blocking liberally and, together with the algorithmic feed, selecting themselves and their followers into a strong filter bubble. I think this is bad and distortionary, but:
It’s not the same censorship, it’s more like people choosing to put blinders on themselves.
On Twitter there’s lots of silent blocking; on LW there is little blocking in general, and when it does happen, historically it has near-universally taken the form of a highly visible rebuke, whether the blocker wants it to or not. I think this bit of culture is unlikely to change regardless of what formal mod tools LW makes available, but the existence of the tools make it slightly easier for people to actually take the strong-public-rebuke action, which is a risky / intellectually brave move, and IMO is warranted more often than it is actually used.
Concretely, if LW got rid of author moderation powers, an author who would otherwise want to block someone could instead reply to a blockworthy-comment with something like:
This is highly escalatory and inflammatory! Lots of onlookers (and the would-be blockee) would strongly contest it and disapprove, almost regardless of the original underlying object-level points. Almost no one who isn’t already extremely well-respected / high-status would ever dare to send something like that. But sometimes it’s still the right thing to say, or at least an accurate and honest reflection of how the author actually feels, and thus worth communicating.
It’s also a way to move the conversation forward (with others, if not the blockee). I view the current mod tools as making this kind of comment marginally easier to make, in a culture where people don’t make it often enough. It’s true that tools and encouragement from the mods or others could move the culture too far in the other direction (cf. Twitter), but I think (based on experience with past mod events, not priors) that is unlikely to happen: author moderation is rarely used and highly visible when it is. Also, the way blocking is actually used in practice on LW, the blocker is often rebuking the community as much or more than the blockee themselves—there’s not much need to block or engage with bad replies that simply get downvoted.
Another reason I think worries about moderation trends are dis-analogous to AI x-risk specifically: moderation is already a big and attention-grabbing problem that lots of people are thinking about and working on in actually-productive ways, and issues with moderation are unlikely to get suddenly and dramatically worse.
In AI x-risk, there’s a common argument (outside of LW) where people will say “Why worry about future hypothetical sci-fi x-risk scenarios when there’s real problems today with algorithmic bias or technological unemployment or whatever”. This is a dumb derail in the context of AI x-risk, but I think it is basically right when it comes to moderation. Why worry about hypothetical future problems and trends when there are so many existing moderation and moderation-adjacent issues warping peoples’ epistemics today? And then you have to look at what people are actually doing about those problems today and what tradeoffs they’re making, and IMO LW having author moderation features seems like a small positive factor (because it can unstick or end unproductive conversations) rather than a large negative one (because it silences critics).
BTW, I’d be up for saying more (or hearing more) with a voice chat if you want. I think “make your case” on LW (the way Zack, Said, etc. have tried in the past) is unlikely to work on me, but I will continue to read the things you post and think about them for myself either way.