The bit­ter les­son for software

16 Mar 2026 23:38 UTC
15 points
3 comments2 min readLW link
(fulcruminc.substack.com)

Types of Hand­off to AIs

Daniel Kokotajlo16 Mar 2026 22:24 UTC
65 points
11 comments8 min readLW link

AICRAFT: DARPA-Funded AI Align­ment Re­searchers — Ap­pli­ca­tions Open

16 Mar 2026 21:44 UTC
67 points
8 comments4 min readLW link

You can’t imi­ta­tion-learn how to con­tinual-learn

Steven Byrnes16 Mar 2026 21:20 UTC
209 points
54 comments6 min readLW link

PSA: Pre­dic­tions mar­kets of­ten have very low liquidity; be care­ful cit­ing them.

Eye You16 Mar 2026 21:07 UTC
126 points
11 comments3 min readLW link

The Plan

Commander Zander16 Mar 2026 20:58 UTC
5 points
0 comments1 min readLW link

What Are My Values?

Corm16 Mar 2026 20:43 UTC
7 points
0 comments8 min readLW link

[Question] Seek­ing Sugges­tions for 2026 S-Pro­cess Recommenders

Ethan Ashkie16 Mar 2026 20:31 UTC
4 points
0 comments1 min readLW link

Car­i­oca Ra­tion­al­ist meetup

Giskard16 Mar 2026 20:30 UTC
2 points
0 comments1 min readLW link

Three Prop­er­ties for Align­ment (and Why We’re Not Train­ing Them)

Quentin FEUILLADE--MONTIXI16 Mar 2026 20:26 UTC
8 points
5 comments3 min readLW link

Do LLMs Have Stable Prefer­ences?

Robert Gambee16 Mar 2026 20:09 UTC
8 points
0 comments7 min readLW link
(github.com)

The Fermi Para­dox Im­plies Domination

Noam Makavy16 Mar 2026 20:04 UTC
1 point
3 comments2 min readLW link

Ad­ding Ty­pos Made Haiku’s Ac­cu­racy Go Up

bira16 Mar 2026 18:31 UTC
31 points
3 comments3 min readLW link

What are the best ways to pub­lish ra­tio­nal fic­tion nowa­days?

Ihor Kendiukhov16 Mar 2026 18:13 UTC
17 points
15 comments3 min readLW link

Compradorization

Benquo16 Mar 2026 16:10 UTC
110 points
11 comments18 min readLW link
(benjaminrosshoffman.com)

SFF-2026 S-Pro­cess Grant Round Ap­pli­ca­tion Announcement

Ethan Ashkie16 Mar 2026 15:57 UTC
18 points
0 comments11 min readLW link
(survivalandflourishing.fund)

Rea­sons to be pes­simistic (and op­ti­mistic) on the fu­ture of biosecurity

Abhishaike Mahajan16 Mar 2026 15:45 UTC
29 points
2 comments41 min readLW link
(www.owlposting.com)

Cus­tomer Satis­fac­tion Opportunities

Tomás B.16 Mar 2026 15:04 UTC
155 points
17 comments13 min readLW link
(open.substack.com)

We found an open weight model that games al­ign­ment honeypots

16 Mar 2026 12:57 UTC
80 points
2 comments10 min readLW link

Monthly Roundup #40: March 2026

Zvi16 Mar 2026 12:20 UTC
30 points
5 comments32 min readLW link
(thezvi.wordpress.com)

Will AI Progress Ac­cel­er­ate or Slow Down? Pro­ject­ing METR Time Horizons

Alvin Ånestrand16 Mar 2026 11:54 UTC
13 points
0 comments10 min readLW link
(forecastingaifutures.substack.com)

Models differ in iden­tity propensities

16 Mar 2026 10:45 UTC
61 points
0 comments14 min readLW link

Ter­rified Com­ments on Cor­rigi­bil­ity in Claude’s Constitution

Zack_M_Davis16 Mar 2026 7:36 UTC
178 points
68 comments11 min readLW link

Hid­den Role Games as a Trusted Model Eval

james.lucassen16 Mar 2026 4:46 UTC
14 points
1 comment9 min readLW link
(jlucassen.com)

Digi­tal Di­chotomy and Why it ex­ists.

Yesh Chala16 Mar 2026 2:02 UTC
8 points
1 comment3 min readLW link

Brown math de­part­ment post­doc­toral position

16 Mar 2026 1:09 UTC
21 points
0 comments1 min readLW link

Hello, World of Mechanis­tic Interpetability

ValueShift Research15 Mar 2026 23:36 UTC
8 points
4 comments5 min readLW link

(I am con­fused about) Non-lin­ear util­i­tar­ian scaling

core15 Mar 2026 23:33 UTC
9 points
6 comments4 min readLW link

Fu­turekind Spring Fel­low­ship 2026 - Ap­pli­ca­tions Now Open

Khushbu Sainani15 Mar 2026 23:28 UTC
2 points
0 comments1 min readLW link

Sched­ule meet­ings us­ing the Pareto principle

beyarkay (Boyd Kane)15 Mar 2026 21:18 UTC
2 points
5 comments2 min readLW link
(boydkane.com)

Was An­thropic that strate­gi­cally in­com­pe­tent?

StanislavKrym15 Mar 2026 20:11 UTC
14 points
0 comments4 min readLW link
(www.lesswrong.com)

What Are We Ac­tu­ally Eval­u­at­ing When We Say a Belief “Tracks Truth”?

Alex Glaucon15 Mar 2026 19:59 UTC
2 points
4 comments6 min readLW link

Su­per­in­tel­li­gence Risk Ed­u­ca­tion that Scales – Lens Academy

15 Mar 2026 19:32 UTC
64 points
2 comments3 min readLW link

Emer­gent stig­mer­gic co­or­di­na­tion in AI agents?

David Africa15 Mar 2026 12:30 UTC
50 points
2 comments3 min readLW link

My Willing Com­plic­ity In “Hu­man Rights Abuse”

AlphaAndOmega15 Mar 2026 10:42 UTC
258 points
37 comments11 min readLW link

Less Ca­pable Misal­igned ASIs Im­ply More Suffering

Ihor Kendiukhov15 Mar 2026 9:36 UTC
11 points
3 comments6 min readLW link

Ra­tion­al­ist Passover Seder in Maryland

Rivka15 Mar 2026 5:42 UTC
3 points
0 comments1 min readLW link

When do in­tu­itions need to be re­li­able?

Anthony DiGiovanni15 Mar 2026 4:18 UTC
8 points
8 comments3 min readLW link

The Ar­tifi­cial Self

15 Mar 2026 1:37 UTC
132 points
13 comments29 min readLW link

Bridge Think­ing and Wall Thinking

Jay Bailey15 Mar 2026 0:20 UTC
48 points
6 comments1 min readLW link

LLM Misal­ign­ment Can be One Gra­di­ent Step Away, and Black­box Eval­u­a­tion Can­not De­tect It.

Yavuz Bakman15 Mar 2026 0:19 UTC
34 points
5 comments3 min readLW link

Walk­ing Math

TickRate15 Mar 2026 0:16 UTC
15 points
2 comments6 min readLW link

How post-train­ing shapes le­gal rep­re­sen­ta­tions: prob­ing SCOTUS opinions across model families

burnssa15 Mar 2026 0:15 UTC
7 points
0 comments8 min readLW link

Self-Recog­ni­tion Fine­tun­ing can Re­v­erse and Prevent Emer­gent Misalignment

15 Mar 2026 0:11 UTC
48 points
24 comments7 min readLW link

Safe AI Ger­many (SAIGE)

Jessica Wang15 Mar 2026 0:10 UTC
6 points
1 comment7 min readLW link

Op­ti­mal (And Eth­i­cal?) Meth­ods To Find “Op­ti­mal Run­ning”

JenniferRM14 Mar 2026 23:16 UTC
9 points
0 comments10 min readLW link

‘Stay­ing with it’ Done Wrong

Selfmaker66214 Mar 2026 22:38 UTC
18 points
0 comments1 min readLW link
(selfmaker.substack.com)

Mini-Mu­nich Suc­ceeds Where KidZa­nia Fails

Novalis14 Mar 2026 22:24 UTC
37 points
0 comments3 min readLW link
(minicities.org)

Fore­cast­ing Dojo Meetup—post­mortem dis­cus­sion.

Vojtech Brynych14 Mar 2026 20:32 UTC
3 points
0 comments1 min readLW link

What con­cerns peo­ple about AI?

spencerg14 Mar 2026 19:24 UTC
34 points
2 comments3 min readLW link
(www.clearerthinking.org)