The Com­pany Man

Tomás B.17 Sep 2025 17:47 UTC
832 points
79 comments18 min readLW link

The Rise of Par­a­sitic AI

Adele Lopez11 Sep 2025 4:38 UTC
771 points
191 comments20 min readLW link

AI 2027: What Su­per­in­tel­li­gence Looks Like

3 Apr 2025 16:23 UTC
697 points
223 comments41 min readLW link
(ai-2027.com)

Eliezer and I wrote a book: If Any­one Builds It, Every­one Dies

So8res14 May 2025 19:00 UTC
655 points
144 comments2 min readLW link

How to Make Superbabies

19 Feb 2025 20:39 UTC
649 points
363 comments31 min readLW link

Ori­ent­ing Toward Wizard Power

johnswentworth8 May 2025 5:23 UTC
592 points
148 comments5 min readLW link

Eliezer’s Un­teach­able Meth­ods of Sanity

Eliezer Yudkowsky7 Dec 2025 2:46 UTC
537 points
157 comments10 min readLW link

A case for courage, when speak­ing of AI danger

So8res27 Jun 2025 2:15 UTC
531 points
130 comments6 min readLW link

What We Learned from Briefing 70+ Law­mak­ers on the Threat from AI

leticiagarcia27 May 2025 18:23 UTC
515 points
17 comments16 min readLW link
(substack.com)

How Does A Blind Model See The Earth?

henry11 Aug 2025 19:58 UTC
501 points
42 comments7 min readLW link
(outsidetext.substack.com)

New En­dorse­ments for “If Any­one Builds It, Every­one Dies”

Malo18 Jun 2025 16:30 UTC
488 points
55 comments4 min readLW link
(intelligence.org)

Ac­countabil­ity Sinks

Martin Sustrik22 Apr 2025 5:00 UTC
462 points
59 comments15 min readLW link
(250bpm.substack.com)

Turn­ing 20 in the prob­a­ble pre-apoc­a­lypse

Parv Mahajan21 Dec 2025 10:14 UTC
459 points
67 comments3 min readLW link

Claude 4.5 Opus’ Soul Document

Richard Weiss28 Nov 2025 23:22 UTC
442 points
44 comments43 min readLW link

The Case Against AI Con­trol Research

johnswentworth21 Jan 2025 16:03 UTC
433 points
85 comments6 min readLW link

HPMOR: The (Prob­a­bly) Un­told Lore

25 Jul 2025 18:39 UTC
432 points
166 comments38 min readLW link

How AI Takeover Might Hap­pen in 2 Years

joshc7 Feb 2025 17:10 UTC
431 points
142 comments29 min readLW link
(x.com)

Will Je­sus Christ re­turn in an elec­tion year?

Eric Neyman24 Mar 2025 16:50 UTC
427 points
59 comments4 min readLW link
(ericneyman.wordpress.com)

the void

nostalgebraist11 Jun 2025 3:19 UTC
427 points
108 comments1 min readLW link
(nostalgebraist.tumblr.com)

Play­ing in the Creek

Hastings10 Apr 2025 17:39 UTC
423 points
13 comments2 min readLW link
(hgreer.com)

Leg­ible vs. Illeg­ible AI Safety Problems

Wei Dai4 Nov 2025 21:39 UTC
399 points
96 comments2 min readLW link

Align­ment re­mains a hard, un­solved problem

evhub27 Nov 2025 8:45 UTC
393 points
98 comments14 min readLW link

AI In­duced Psy­chosis: A shal­low investigation

Tim Hua26 Aug 2025 20:03 UTC
390 points
47 comments27 min readLW link

Hospi­tal­iza­tion: A Review

Logan Riggs9 Oct 2025 14:36 UTC
380 points
21 comments9 min readLW link

A deep cri­tique of AI 2027’s bad timeline models

titotal19 Jun 2025 13:29 UTC
378 points
40 comments39 min readLW link
(titotal.substack.com)

6 rea­sons why “al­ign­ment-is-hard” dis­course seems alien to hu­man in­tu­itions, and vice-versa

Steven Byrnes3 Dec 2025 18:37 UTC
377 points
92 comments17 min readLW link

A Bear Case: My Pre­dic­tions Re­gard­ing AI Progress

Thane Ruthenis5 Mar 2025 16:41 UTC
376 points
167 comments9 min readLW link

Gen­er­al­ized Han­gri­ness: A Stan­dard Ra­tion­al­ist Stance Toward Emotions

johnswentworth10 Jul 2025 18:22 UTC
374 points
71 comments7 min readLW link

What’s the short timeline plan?

Marius Hobbhahn2 Jan 2025 14:59 UTC
373 points
51 comments23 min readLW link

Four ways learn­ing Econ makes peo­ple dumber re: fu­ture AI

Steven Byrnes21 Aug 2025 17:52 UTC
370 points
52 comments6 min readLW link
(x.com)

VDT: a solu­tion to de­ci­sion theory

L Rudolf L1 Apr 2025 21:04 UTC
370 points
34 comments4 min readLW link

LessWrong has been ac­quired by EA

habryka1 Apr 2025 13:09 UTC
366 points
55 comments1 min readLW link

Para­noia: A Begin­ner’s Guide

habryka13 Nov 2025 7:56 UTC
363 points
70 comments13 min readLW link

Re­cent AI model progress feels mostly like bullshit

lc24 Mar 2025 19:28 UTC
362 points
89 comments8 min readLW link
(zeropath.com)

Sublimi­nal Learn­ing: LLMs Trans­mit Be­hav­ioral Traits via Hid­den Sig­nals in Data

22 Jul 2025 16:37 UTC
348 points
40 comments4 min readLW link

Global Call for AI Red Lines—Signed by No­bel Lau­re­ates, Former Heads of State, and 200+ Promi­nent Figures

Charbel-Raphaël22 Sep 2025 18:22 UTC
345 points
27 comments6 min readLW link

In­ter­pretabil­ity Will Not Reli­ably Find De­cep­tive AI

Neel Nanda4 May 2025 16:32 UTC
343 points
69 comments7 min readLW link

Policy for LLM Writ­ing on LessWrong

24 Mar 2025 21:41 UTC
343 points
72 comments2 min readLW link

Why I Tran­si­tioned: A Case Study

Fiora Starlight1 Nov 2025 22:58 UTC
341 points
82 comments10 min readLW link

Emer­gent Misal­ign­ment: Nar­row fine­tun­ing can pro­duce broadly mis­al­igned LLMs

25 Feb 2025 17:39 UTC
335 points
92 comments4 min readLW link

The Problem

5 Aug 2025 21:40 UTC
331 points
220 comments26 min readLW link

I ate bear fat with honey and salt flakes, to prove a point

aggliu4 Nov 2025 2:00 UTC
330 points
53 comments5 min readLW link
(signoregalilei.com)

So You Think You’ve Awo­ken ChatGPT

JustisMills11 Jul 2025 1:01 UTC
329 points
88 comments9 min readLW link

Why you should eat meat—even if you hate fac­tory farming

KatWoods25 Sep 2025 15:39 UTC
326 points
101 comments10 min readLW link

Be­ware Gen­eral Claims about “Gen­er­al­iz­able Rea­son­ing Ca­pa­bil­ities” (of Modern AI Sys­tems)

LawrenceC11 Jun 2025 19:27 UTC
318 points
19 comments16 min readLW link

Toss a bit­coin to your Light­cone – LW + Lighthaven’s 2026 fundraiser

habryka13 Dec 2025 19:32 UTC
316 points
129 comments52 min readLW link

Make More Grayspaces

Duncan Sabien (Inactive)19 Jul 2025 22:22 UTC
315 points
65 comments13 min readLW link

Towards a Ty­pol­ogy of Strange LLM Chains-of-Thought

1a3orn9 Oct 2025 22:02 UTC
311 points
29 comments9 min readLW link

So You Want To Make Marginal Progress...

johnswentworth7 Feb 2025 23:22 UTC
311 points
42 comments4 min readLW link

Trac­ing the Thoughts of a Large Lan­guage Model

Adam Jermyn27 Mar 2025 17:20 UTC
308 points
23 comments10 min readLW link
(www.anthropic.com)