The Terrarium

Caleb Biddulph26 Mar 2026 18:08 UTC
611 points
54 comments21 min readLW link

Less Dead

Aurelia11 Mar 2026 5:07 UTC
555 points
140 comments8 min readLW link

On The In­de­pen­dence Axiom

Ihor Kendiukhov8 Mar 2026 14:38 UTC
336 points
93 comments23 min readLW link

My hobby: run­ning de­ranged surveys

leogao27 Mar 2026 0:41 UTC
321 points
66 comments9 min readLW link

Socrates is Mortal

Benquo26 Mar 2026 15:34 UTC
288 points
57 comments10 min readLW link
(benjaminrosshoffman.com)

Re­quiem for a Tran­shu­man Timeline

Ihor Kendiukhov17 Mar 2026 21:27 UTC
287 points
47 comments5 min readLW link

Gemma Needs Help

annajs10 Mar 2026 17:39 UTC
284 points
22 comments6 min readLW link

An Align­ment Jour­nal: Com­ing Soon

3 Mar 2026 20:27 UTC
271 points
34 comments6 min readLW link
(blog.alignmentjournal.org)

My Willing Com­plic­ity In “Hu­man Rights Abuse”

AlphaAndOmega15 Mar 2026 10:42 UTC
258 points
37 comments11 min readLW link

No, we haven’t up­loaded a fly yet

Ariel Zeleznikow-Johnston19 Mar 2026 23:43 UTC
243 points
8 comments8 min readLW link
(open.substack.com)

Is fever a symp­tom of glycine defi­ciency?

Benquo22 Mar 2026 14:44 UTC
218 points
86 comments6 min readLW link
(benjaminrosshoffman.com)

You can’t imi­ta­tion-learn how to con­tinual-learn

Steven Byrnes16 Mar 2026 21:20 UTC
209 points
54 comments6 min readLW link

Maybe there’s a pat­tern here?

dynomight4 Mar 2026 20:32 UTC
193 points
45 comments7 min readLW link

Broad Timelines

Toby_Ord19 Mar 2026 19:05 UTC
191 points
25 comments16 min readLW link

“The AI Doc” is com­ing out March 26

19 Mar 2026 22:55 UTC
189 points
2 comments1 min readLW link

Don’t Let LLMs Write For You

JustisMills10 Mar 2026 18:49 UTC
184 points
45 comments3 min readLW link
(justismills.substack.com)

Nul­lius in Verba: 3rd party ev­i­dence for Nec­tome’s Brain Preservation

Aurelia20 Mar 2026 3:19 UTC
179 points
18 comments12 min readLW link

Ter­rified Com­ments on Cor­rigi­bil­ity in Claude’s Constitution

Zack_M_Davis16 Mar 2026 7:36 UTC
178 points
68 comments11 min readLW link

So­lar Storms

Croissanthology8 Mar 2026 14:04 UTC
173 points
43 comments12 min readLW link

Per­son­al­ity Self-Replicators

eggsyntax5 Mar 2026 20:30 UTC
173 points
51 comments10 min readLW link

Eco­nomic effi­ciency of­ten un­der­mines so­ciopoli­ti­cal autonomy

Richard_Ngo10 Mar 2026 19:30 UTC
166 points
38 comments12 min readLW link
(www.mindthefuture.info)

Prologue to Ter­rified Com­ments on Claude’s Constitution

Zack_M_Davis9 Mar 2026 6:46 UTC
165 points
27 comments8 min readLW link
(zackmdavis.net)

Some things I no­ticed while LARPing as a grantmaker

Zach Stein-Perlman23 Mar 2026 15:00 UTC
163 points
11 comments7 min readLW link

My Most Costly Delusion

Ihor Kendiukhov22 Mar 2026 12:21 UTC
160 points
14 comments3 min readLW link

Cus­tomer Satis­fac­tion Opportunities

Tomás B.16 Mar 2026 15:04 UTC
155 points
17 comments13 min readLW link
(open.substack.com)

OpenAI’s surveillance lan­guage has many po­ten­tial loop­holes and they can do better

Tom Smith4 Mar 2026 4:25 UTC
153 points
4 comments10 min readLW link

The Case for Low-Com­pe­tence ASI Failure Scenarios

Ihor Kendiukhov19 Mar 2026 23:10 UTC
146 points
8 comments6 min readLW link

(Some) Nat­u­ral Emer­gent Misal­ign­ment from Re­ward Hack­ing in Non-Pro­duc­tion RL

30 Mar 2026 10:56 UTC
144 points
8 comments18 min readLW link

Thoughts on the Pause AI protest

philh6 Mar 2026 21:50 UTC
136 points
18 comments7 min readLW link
(reasonableapproximation.net)

The Ar­tifi­cial Self

15 Mar 2026 1:37 UTC
132 points
13 comments29 min readLW link

Dis­patch from An­thropic v. Depart­ment of War Pre­limi­nary In­junc­tion Mo­tion Hearing

Zack_M_Davis26 Mar 2026 2:54 UTC
130 points
13 comments7 min readLW link

New LessWrong Edi­tor! (Also, an up­date to our LLM policy.)

RobertM14 Mar 2026 3:33 UTC
128 points
130 comments5 min readLW link

PSA: Pre­dic­tions mar­kets of­ten have very low liquidity; be care­ful cit­ing them.

Eye You16 Mar 2026 21:07 UTC
126 points
11 comments3 min readLW link

China Derange­ment Syndrome

Arjun Panickssery21 Mar 2026 19:19 UTC
124 points
70 comments4 min readLW link
(arjunpanickssery.substack.com)

Physics of RL: Toy scal­ing laws for the emer­gence of re­ward-seeking

Alex Meinke4 Mar 2026 8:12 UTC
120 points
9 comments10 min readLW link

The case for sa­ti­at­ing cheaply-satis­fied AI preferences

Alex Mallen10 Mar 2026 18:09 UTC
112 points
7 comments23 min readLW link

Stan­ley Mil­gram wasn’t pes­simistic enough about hu­man na­ture?

David Gross28 Mar 2026 14:22 UTC
111 points
16 comments3 min readLW link

The cur­rent SOTA model was re­leased with­out safety evals

8 Mar 2026 1:51 UTC
111 points
12 comments5 min readLW link

Compradorization

Benquo16 Mar 2026 16:10 UTC
110 points
11 comments18 min readLW link
(benjaminrosshoffman.com)

The Lethal Real­ity Hypothesis

Ihor Kendiukhov11 Mar 2026 15:23 UTC
109 points
25 comments20 min readLW link

How well do mod­els fol­low their con­sti­tu­tions?

12 Mar 2026 0:07 UTC
107 points
5 comments26 min readLW link

Game Rec­og­nizes Game

eva_3 Mar 2026 10:09 UTC
105 points
15 comments12 min readLW link

The Distaff Texts

Tomás B.20 Mar 2026 15:05 UTC
105 points
6 comments14 min readLW link

Folie à Ma­chine: LLMs and Epistemic Capture

DaystarEld29 Mar 2026 15:23 UTC
104 points
21 comments21 min readLW link

Why AI Eval­u­a­tion Regimes are bad

12 Mar 2026 13:59 UTC
102 points
12 comments9 min readLW link
(cognition.cafe)

Let­ting Claude do Au­tonomous Re­search to Im­prove SAEs

chanind10 Mar 2026 18:52 UTC
102 points
16 comments7 min readLW link

The state of AI safety in four fake graphs

Boaz Barak30 Mar 2026 13:21 UTC
102 points
39 comments2 min readLW link

The Elect

Tomás B.6 Mar 2026 15:34 UTC
100 points
1 comment16 min readLW link
(open.substack.com)

Oper­a­tional­iz­ing FDT

Vivek Hebbar13 Mar 2026 0:12 UTC
99 points
11 comments6 min readLW link

LLMs as Gi­ant Lookup-Tables of Shal­low Circuits

17 Mar 2026 21:35 UTC
99 points
35 comments7 min readLW link