In­tro­duc­ing the AE Align­ment Pod­cast (Ep. 1: En­doge­nous Steer­ing Re­sis­tance with Alex McKen­zie)

27 Mar 2026 22:13 UTC
24 points
3 comments1 min readLW link

What if the US loses the 2026 Hor­muz Conflict

timothy liptrot27 Mar 2026 21:03 UTC
23 points
2 comments6 min readLW link

AI Safety Guide for TRUE Begin­ners by TRUE begginers

27 Mar 2026 20:38 UTC
11 points
0 comments6 min readLW link
(forum.effectivealtruism.org)

Pray for Casanova

Tomás B.27 Mar 2026 20:24 UTC
52 points
1 comment5 min readLW link

SB 53 and RAISE im­ple­men­ta­tion roles

Eric Neyman27 Mar 2026 20:04 UTC
22 points
0 comments2 min readLW link

Con­crete pro­jects to pre­pare for superintelligence

27 Mar 2026 20:04 UTC
23 points
0 comments11 min readLW link
(www.forethought.org)

Mi­ni­a­ture Cities Might Be the Non-Co­er­cive Schools Many Thought Were Impossible

Novalis27 Mar 2026 18:47 UTC
9 points
2 comments11 min readLW link
(minicities.org)

Con­trolAI 2025 Im­pact Re­port: our progress to­ward an in­ter­na­tional ban on ASI

27 Mar 2026 18:10 UTC
75 points
4 comments4 min readLW link
(controlai.com)

AI’s ca­pa­bil­ity im­prove­ments haven’t come from it get­ting less affordable

Anders Cairns Woodruff27 Mar 2026 17:09 UTC
87 points
0 comments6 min readLW link

Launch­ing Eu­zoia: Bee­minder for Effec­tive Charities

27 Mar 2026 15:56 UTC
9 points
0 comments1 min readLW link
(app.euzoia.org)

Stop ask­ing “how good is this” to de­cide be­tween dona­tion op­por­tu­ni­ties I recommend

Zach Stein-Perlman27 Mar 2026 15:00 UTC
35 points
6 comments1 min readLW link

Startup Les­sons for AI Safety

27 Mar 2026 14:35 UTC
18 points
3 comments3 min readLW link

An­thropic vs. DoW #6: The Court Rules

Zvi27 Mar 2026 11:40 UTC
45 points
2 comments31 min readLW link
(thezvi.wordpress.com)

A Tax­on­omy of Agents: In­tro & Re­quest for feedback

Jonas Hallgren27 Mar 2026 10:03 UTC
13 points
4 comments3 min readLW link
(equilibria1.substack.com)

Why Mo­ral Ques­tions Get De­cided, Not Answered

Alex Glaucon27 Mar 2026 8:33 UTC
12 points
9 comments9 min readLW link

COT con­trol: The Word Dis­ap­pears, but the Thought Does Not

Pranjal Garg27 Mar 2026 6:15 UTC
10 points
10 comments10 min readLW link

Are we al­ign­ing the model or just its mask?

James Sullivan27 Mar 2026 2:10 UTC
11 points
0 comments10 min readLW link
(substack.com)

One World Govern­ment by 2150

Julius27 Mar 2026 1:19 UTC
11 points
4 comments3 min readLW link

An­a­lyz­ing the claim “The most fun­da­men­tal right is the right to ex­ist”

Akseli Jussinmäki27 Mar 2026 0:47 UTC
8 points
3 comments3 min readLW link

My hobby: run­ning de­ranged surveys

leogao27 Mar 2026 0:41 UTC
321 points
66 comments9 min readLW link

Pre­limi­nary Re­sults on Build­ing Graphs from SAEs

ZachMaas27 Mar 2026 0:39 UTC
13 points
0 comments5 min readLW link
(zachmaas.com)

Un­der­stand­ing and track­ing de­vel­op­ments in robotics

janvi26 Mar 2026 23:51 UTC
6 points
0 comments5 min readLW link

Scaf­folded Re­pro­duc­ers, Scaf­folded Agents

Mateusz Bagiński26 Mar 2026 23:47 UTC
37 points
2 comments3 min readLW link

Test your best meth­ods on our hard CoT in­terp tasks

26 Mar 2026 19:24 UTC
59 points
2 comments19 min readLW link

The Terrarium

Caleb Biddulph26 Mar 2026 18:08 UTC
611 points
54 comments21 min readLW link

What if su­per­in­tel­li­gence is just weak?

Simon Lermen26 Mar 2026 17:45 UTC
33 points
25 comments2 min readLW link
(substack.com)

The con­tin­u­ous tense is dis­ap­pear­ing from your life

PatrickDFarley26 Mar 2026 17:14 UTC
10 points
0 comments5 min readLW link

“What Ex­actly Would An In­ter­na­tional AI Treaty Say?” Is a Bad Objection

Davidmanheim26 Mar 2026 16:29 UTC
92 points
11 comments6 min readLW link

Socrates is Mortal

Benquo26 Mar 2026 15:34 UTC
288 points
57 comments10 min readLW link
(benjaminrosshoffman.com)

Five years since lockdown

mingyuan26 Mar 2026 15:29 UTC
37 points
1 comment4 min readLW link
(mingyuan.substack.com)

You can just mul­ti­ply point es­ti­mates (if you only care about EV)

Zach Stein-Perlman26 Mar 2026 15:00 UTC
29 points
9 comments2 min readLW link

AI #161 Part 1: 80,000 Interviews

Zvi26 Mar 2026 13:20 UTC
33 points
1 comment28 min readLW link
(thezvi.wordpress.com)

Sen. San­ders (I-VT) and Rep. Oca­sio-Cortez (D-NY) pro­pose AI Data Cen­ter Mo­ra­to­rium Act

Matrice Jacobine26 Mar 2026 13:13 UTC
40 points
1 comment1 min readLW link

Past Au­toma­tion Re­placed Jobs. AI Will Re­place Work­ers.

James_Miller26 Mar 2026 12:32 UTC
52 points
6 comments13 min readLW link

Fine Tun­ing CoT obfus­ca­tion into Kimi K2.5

Graeme Ford26 Mar 2026 4:02 UTC
15 points
2 comments11 min readLW link

Dis­patch from An­thropic v. Depart­ment of War Pre­limi­nary In­junc­tion Mo­tion Hearing

Zack_M_Davis26 Mar 2026 2:54 UTC
130 points
13 comments7 min readLW link

La­bel By Us­able Volume

jefftk26 Mar 2026 2:30 UTC
43 points
12 comments1 min readLW link
(www.jefftk.com)

Who’s Afraid of Acausal Trades?

edgecase6426 Mar 2026 1:53 UTC
15 points
8 comments9 min readLW link

A Black Box Made Less Opaque (part 3)

Matthew McDonnell26 Mar 2026 1:41 UTC
7 points
0 comments18 min readLW link

Mo­ral Ex­ten­sion Risk

Tentrion26 Mar 2026 0:21 UTC
10 points
1 comment4 min readLW link

How do you eval­u­ate AI ca­pa­bil­ity claims in ac­tual soft­ware prod­ucts?

Dhruv Gulati26 Mar 2026 0:18 UTC
6 points
1 comment1 min readLW link

Bidi­rec­tion­al­ity is the Ob­vi­ous BCI Paradigm

Elliot Callender25 Mar 2026 22:07 UTC
12 points
0 comments3 min readLW link

PauseAI Capi­tol Day of Action

maia25 Mar 2026 21:21 UTC
9 points
0 comments1 min readLW link

A Toy En­vi­ron­ment For Ex­plor­ing Rea­son­ing About Reward

25 Mar 2026 20:29 UTC
60 points
10 comments3 min readLW link

Im­mor­tal­ity: A Beg­giner’s Guide (Part 3)

MarkelKori25 Mar 2026 19:10 UTC
4 points
0 comments4 min readLW link

Find­ing X-Risks and S-Risks by Gra­di­ent Descent

dspeyer25 Mar 2026 17:58 UTC
5 points
2 comments2 min readLW link

The Scary Bridge

moridinamael25 Mar 2026 17:11 UTC
16 points
0 comments2 min readLW link

How to do illu­sion­ist meditation

Jack Thompson25 Mar 2026 17:10 UTC
9 points
0 comments11 min readLW link

Can Agents Fool Each Other? Find­ings from the AI Village

Shoshannah Tekofsky25 Mar 2026 17:05 UTC
71 points
7 comments3 min readLW link
(theaidigest.org)

I lost my faith in in­tro­spec­tion—and you can too!

Jack Thompson25 Mar 2026 17:00 UTC
9 points
8 comments4 min readLW link