AI 2027 ver­sus World War 2027

Mitchell_Porter24 Mar 2026 23:57 UTC
9 points
6 comments2 min readLW link

Book Re­view: Open Socrates (Part 1)

Zvi24 Mar 2026 22:21 UTC
32 points
4 comments116 min readLW link
(thezvi.wordpress.com)

Book Re­view: Open Socrates (Part 2)

Zvi24 Mar 2026 22:20 UTC
22 points
2 comments81 min readLW link
(thezvi.wordpress.com)

Agents Can Get Stuck in Self-dis­trust­ing Equilibria

Ashe Vazquez Nuñez24 Mar 2026 22:05 UTC
33 points
2 comments12 min readLW link

La­tent In­tro­spec­tion (and other open-source in­tro­spec­tion pa­pers)

24 Mar 2026 21:23 UTC
98 points
3 comments9 min readLW link
(arxiv.org)

An In­for­mal Defi­ni­tion of Goals for Embed­ded Agents

Ashe Vazquez Nuñez24 Mar 2026 18:36 UTC
14 points
0 comments1 min readLW link

My cost-effec­tive­ness unit

Zach Stein-Perlman24 Mar 2026 15:30 UTC
65 points
5 comments4 min readLW link

AI Safety Newslet­ter #70: Au­to­mated War­fare and AI Layoffs

24 Mar 2026 15:30 UTC
8 points
0 comments4 min readLW link
(newsletter.safe.ai)

Mon­day AI Radar #18

Against Moloch24 Mar 2026 15:15 UTC
7 points
4 comments8 min readLW link
(againstmoloch.com)

The Fourth World

Linch24 Mar 2026 13:43 UTC
27 points
16 comments6 min readLW link

Safe Re­cur­sive Self-Im­prove­ment with Ver­ified Compilers

Adam Chlipala24 Mar 2026 13:35 UTC
15 points
0 comments11 min readLW link

Com­par­ing Across Pos­si­ble Worlds

unruly abstractions24 Mar 2026 10:09 UTC
7 points
4 comments5 min readLW link

Com­ing of Age: Chap­ters 1 and 2

Ihor Kendiukhov24 Mar 2026 9:15 UTC
13 points
0 comments1 min readLW link

The AIXI per­spec­tive on AI Safety

Cole Wyeth24 Mar 2026 3:24 UTC
80 points
4 comments6 min readLW link

Con­tra Dances Should Avoid Saturdays

jefftk24 Mar 2026 2:30 UTC
11 points
0 comments1 min readLW link
(www.jefftk.com)

Malmö AI Safety meetup

Vadym Sulzhenko (Vaigotaku)24 Mar 2026 2:29 UTC
1 point
0 comments1 min readLW link

In­for­ma­tion Overdose

Tridiv Sharma24 Mar 2026 2:26 UTC
−2 points
0 comments3 min readLW link

We can­not safely au­to­mate value al­ign­ment eval­u­a­tion and re­search with­out think­ing about del­e­ga­tion and discretion

Mflena24 Mar 2026 2:25 UTC
8 points
0 comments10 min readLW link

Every Ma­jor LLM is a 1-Box Smok­ing Thirder

Olivia Scharfman24 Mar 2026 2:18 UTC
14 points
4 comments10 min readLW link

A ToM-In­spired Agenda for AI Safety Research

Andrés Cotton24 Mar 2026 2:13 UTC
7 points
1 comment5 min readLW link

Pok­ing and Edit­ing the Circuits

unruly abstractions24 Mar 2026 1:15 UTC
2 points
0 comments8 min readLW link

Adopt a de­bug­ger’s mind­set to solve your re­cur­ring life problems

Declan Molony24 Mar 2026 0:22 UTC
24 points
6 comments8 min readLW link

Hol­ly­wood ACX Meetup—April 2026

Timothy M.23 Mar 2026 23:55 UTC
1 point
0 comments1 min readLW link

Ex­per­i­men­tal Ev­i­dence for Si­mu­la­tor The­ory— Part 2: The Scalers Strike Back

RogerDearnaley23 Mar 2026 22:37 UTC
21 points
0 comments34 min readLW link

Ex­per­i­men­tal Ev­i­dence for Si­mu­la­tor The­ory— Part 1: Emer­gent Misal­ign­ment and Weird Generalizations

RogerDearnaley23 Mar 2026 22:37 UTC
25 points
0 comments53 min readLW link

How Far Can Ob­ser­va­tion Take Us?

unruly abstractions23 Mar 2026 21:56 UTC
12 points
0 comments9 min readLW link

Vibe­coders can’t build for longevity

dominicq23 Mar 2026 19:36 UTC
13 points
9 comments4 min readLW link

Ablat­ing Split Per­son­al­ity Training

OscarGilg23 Mar 2026 17:45 UTC
55 points
1 comment5 min readLW link

Mea­sur­ing and im­prov­ing cod­ing au­dit re­al­ism with de­ploy­ment resources

23 Mar 2026 17:20 UTC
43 points
1 comment10 min readLW link
(alignment.anthropic.com)

AI char­ac­ter is a big deal

23 Mar 2026 16:36 UTC
34 points
33 comments12 min readLW link
(www.forethought.org)

Some things I no­ticed while LARPing as a grantmaker

Zach Stein-Perlman23 Mar 2026 15:00 UTC
163 points
11 comments7 min readLW link

Rep­re­sen­ta­tive Futarchy

goldfine23 Mar 2026 14:39 UTC
2 points
0 comments2 min readLW link
(itsnotgambling.substack.com)

Which types of AI al­ign­ment re­search are most likely to be good for all sen­tient be­ings?

MichaelDickens23 Mar 2026 13:38 UTC
11 points
1 comment6 min readLW link

Kelly Cri­te­rion is for Cowards

X4vier23 Mar 2026 1:56 UTC
16 points
28 comments3 min readLW link

Set the Line Be­fore It’s Crossed

nomagicpill23 Mar 2026 1:25 UTC
12 points
0 comments5 min readLW link
(nomagicpill.substack.com)

When Align­ment Be­comes an At­tack Sur­face: Prompt In­jec­tion in Co­op­er­a­tive Multi-Agent Systems

Cornelis Dirk Haupt23 Mar 2026 0:22 UTC
9 points
0 comments4 min readLW link

At­tend the 2026 Re­pro­duc­tive Fron­tiers Sum­mit, June 16–18, Berkeley

22 Mar 2026 21:15 UTC
97 points
1 comment5 min readLW link

You’re ab­solutely right, Se­na­tor. I was be­ing naive about the poli­ti­cal re­al­ity.

Chris Datcu22 Mar 2026 20:42 UTC
4 points
5 comments2 min readLW link

Let’s Rea­son About (Your) Job Se­cu­rity!

Gergely Máté22 Mar 2026 20:13 UTC
8 points
0 comments4 min readLW link

Madi­son ACX Mee­tups Every­where

edalva22 Mar 2026 17:22 UTC
1 point
0 comments1 min readLW link

Is fever a symp­tom of glycine defi­ciency?

Benquo22 Mar 2026 14:44 UTC
218 points
86 comments6 min readLW link
(benjaminrosshoffman.com)

My Most Costly Delusion

Ihor Kendiukhov22 Mar 2026 12:21 UTC
161 points
14 comments3 min readLW link

Notic­ing a Teacher’s Pass­word Pattern

Dentosal22 Mar 2026 9:10 UTC
26 points
12 comments2 min readLW link

The Toy Story Saga is not yet finished

Raemon22 Mar 2026 3:53 UTC
95 points
9 comments8 min readLW link

Key to Life No. 9: Access

MarkelKori21 Mar 2026 21:53 UTC
11 points
0 comments3 min readLW link

My Ham­mer­time Fi­nal Exam

evjeny21 Mar 2026 20:40 UTC
11 points
0 comments2 min readLW link

Un­der­stand­ing when and why agents scheme

21 Mar 2026 20:33 UTC
50 points
2 comments4 min readLW link

Build­ing a Web App Us­ing an AI-As­sisted Workflow

Thomas Castriensis21 Mar 2026 19:21 UTC
1 point
0 comments1 min readLW link

China Derange­ment Syndrome

Arjun Panickssery21 Mar 2026 19:19 UTC
124 points
70 comments4 min readLW link
(arjunpanickssery.substack.com)

China de­clares AGI de­vel­op­ment to be a part of 5-year plan

Darmani21 Mar 2026 17:21 UTC
32 points
4 comments1 min readLW link