Take note of how bright­ness makes you feel

Biff Wiff30 Mar 2026 23:48 UTC
47 points
6 comments3 min readLW link

Cost of Cul­tured Meat: work­shop, mod­el­ing, re­sources, feedback

david reinstein30 Mar 2026 23:44 UTC
3 points
0 comments4 min readLW link

Pan­gram (AI de­tec­tion soft­ware) can be evaded

Eye You30 Mar 2026 23:21 UTC
23 points
11 comments3 min readLW link

A Mir­ror Test For LLMs

Christopher Ackerman30 Mar 2026 22:44 UTC
14 points
0 comments24 min readLW link

God Can Send An Email

AlphaAndOmega30 Mar 2026 22:36 UTC
17 points
12 comments9 min readLW link

On Bad­ness of Death

MarkelKori30 Mar 2026 20:34 UTC
3 points
3 comments3 min readLW link

How to Solve Se­cure Pro­gram Synthesis

30 Mar 2026 20:12 UTC
24 points
0 comments11 min readLW link

Block­ing live failures with syn­chronous monitors

30 Mar 2026 17:44 UTC
24 points
0 comments4 min readLW link
(open.substack.com)

A Guide to the The­ory of Ap­pro­pri­ate­ness Papers

Joel Z. Leibo30 Mar 2026 16:56 UTC
3 points
1 comment1 min readLW link

AI should be a good cit­i­zen, not just a good assistant

30 Mar 2026 14:32 UTC
37 points
16 comments9 min readLW link
(www.forethought.org)

My One-Year-Old Pre­dic­tions for What the World Will Look Like in 3 Years

Ihor Kendiukhov30 Mar 2026 13:56 UTC
12 points
2 comments3 min readLW link

Propo­si­tional Alignment

williawa30 Mar 2026 13:50 UTC
14 points
0 comments2 min readLW link

AI #161 Part 2: Every De­bate on AI

Zvi30 Mar 2026 13:40 UTC
28 points
5 comments36 min readLW link
(thezvi.wordpress.com)

The state of AI safety in four fake graphs

Boaz Barak30 Mar 2026 13:21 UTC
102 points
39 comments2 min readLW link

(Some) Nat­u­ral Emer­gent Misal­ign­ment from Re­ward Hack­ing in Non-Pro­duc­tion RL

30 Mar 2026 10:56 UTC
144 points
9 comments18 min readLW link

Why Cor­rigi­bil­ity Mat­ters (If It Mat­ters At All)

Savannah Harlan30 Mar 2026 10:11 UTC
10 points
5 comments13 min readLW link

I am definitely miss­ing the pre-AI writ­ing era

N. Cailie29 Mar 2026 19:10 UTC
41 points
7 comments2 min readLW link

Claude’s con­sti­tu­tion is great

Oscar29 Mar 2026 17:53 UTC
13 points
2 comments1 min readLW link

Claude has no baseline

Dave92F129 Mar 2026 17:15 UTC
−8 points
0 comments1 min readLW link
(mugwumpery.com)

Folie à Ma­chine: LLMs and Epistemic Capture

DaystarEld29 Mar 2026 15:23 UTC
104 points
21 comments21 min readLW link

The Power of Assumption

Ysamuels29 Mar 2026 11:58 UTC
15 points
4 comments2 min readLW link

Park­in­son’s Law of Worry

Jakub Halmeš29 Mar 2026 11:50 UTC
41 points
6 comments1 min readLW link
(unpredictabletokens.substack.com)

“Path to Vic­tory”

Chris_Leong29 Mar 2026 6:23 UTC
26 points
4 comments6 min readLW link

Track­ing (Ex­pert/​In­fluen­tial) Pre­dic­tions about AI

Noah Birnbaum28 Mar 2026 23:10 UTC
25 points
1 comment2 min readLW link

The Skill of Us­ing AI Agents Well

becausecurious28 Mar 2026 22:57 UTC
20 points
6 comments6 min readLW link

Heed­ful­ness Workouts

Thomas Castriensis28 Mar 2026 21:00 UTC
15 points
2 comments3 min readLW link
(thomascastriensis.substack.com)

[Story] Hu­man Align­ment Isn’t Enough

pku28 Mar 2026 18:38 UTC
20 points
0 comments4 min readLW link

Don’t Over­dose Lo­cally Benefi­cial Changes

Mateusz Bagiński28 Mar 2026 18:24 UTC
81 points
12 comments4 min readLW link

Com­pre­hen­sive FAQ on Immortalism

MarkelKori28 Mar 2026 16:03 UTC
7 points
2 comments1 min readLW link

Nick Bostrom: How big is the cos­mic en­dow­ment?

Zach Stein-Perlman28 Mar 2026 15:00 UTC
72 points
18 comments3 min readLW link
(amazon.com)

In­fu­sion: Re­v­erse en­g­ineer­ing In­fluence Func­tions to craft train­ing doc­u­ments...

J Rosser28 Mar 2026 14:40 UTC
5 points
0 comments2 min readLW link

Stan­ley Mil­gram wasn’t pes­simistic enough about hu­man na­ture?

David Gross28 Mar 2026 14:22 UTC
111 points
16 comments3 min readLW link

How to Do the Mar­quette Method, a Ba­sic Guide (cross­post)

Psmith28 Mar 2026 13:55 UTC
4 points
3 comments4 min readLW link

The Prob­lem with Ask­ing your Doctor

ChristianKl28 Mar 2026 11:14 UTC
25 points
29 comments1 min readLW link

Would a con­struc­tive proof of de­ter­minism be use­ful?

Tridiv Sharma28 Mar 2026 11:01 UTC
−9 points
6 comments2 min readLW link

Just Use Bayes: Sleep­ing Beauty and Monty Hall

Steff28 Mar 2026 2:59 UTC
12 points
37 comments12 min readLW link

What Makes a Good Ter­mi­nal Bench Task

Ivan Bercovich28 Mar 2026 2:26 UTC
12 points
0 comments12 min readLW link

An­thropic vs. DoW Pre­limi­nary In­junc­tion Ruling

anaguma28 Mar 2026 1:52 UTC
12 points
0 comments48 min readLW link

Why should I have opinions about AI timelines?

cjiang28 Mar 2026 0:19 UTC
6 points
0 comments4 min readLW link

Do fron­tier LLMs still ex­press differ­ent val­ues in differ­ent lan­guages?

Ibrahim Ahmed28 Mar 2026 0:05 UTC
19 points
0 comments1 min readLW link

In­tro­duc­ing the AE Align­ment Pod­cast (Ep. 1: En­doge­nous Steer­ing Re­sis­tance with Alex McKen­zie)

27 Mar 2026 22:13 UTC
24 points
3 comments1 min readLW link

What if the US loses the 2026 Hor­muz Conflict

timothy liptrot27 Mar 2026 21:03 UTC
23 points
2 comments6 min readLW link

AI Safety Guide for TRUE Begin­ners by TRUE begginers

27 Mar 2026 20:38 UTC
11 points
0 comments6 min readLW link
(forum.effectivealtruism.org)

Pray for Casanova

Tomás B.27 Mar 2026 20:24 UTC
52 points
1 comment5 min readLW link

SB 53 and RAISE im­ple­men­ta­tion roles

Eric Neyman27 Mar 2026 20:04 UTC
22 points
0 comments2 min readLW link

Con­crete pro­jects to pre­pare for superintelligence

27 Mar 2026 20:04 UTC
23 points
0 comments11 min readLW link
(www.forethought.org)

Mi­ni­a­ture Cities Might Be the Non-Co­er­cive Schools Many Thought Were Impossible

Novalis27 Mar 2026 18:47 UTC
9 points
2 comments11 min readLW link
(minicities.org)

Con­trolAI 2025 Im­pact Re­port: our progress to­ward an in­ter­na­tional ban on ASI

27 Mar 2026 18:10 UTC
75 points
4 comments4 min readLW link
(controlai.com)

AI’s ca­pa­bil­ity im­prove­ments haven’t come from it get­ting less affordable

Anders Cairns Woodruff27 Mar 2026 17:09 UTC
87 points
0 comments6 min readLW link

Launch­ing Eu­zoia: Bee­minder for Effec­tive Charities

27 Mar 2026 15:56 UTC
9 points
0 comments1 min readLW link
(app.euzoia.org)