Bidi­rec­tion­al­ity is the Ob­vi­ous BCI Paradigm

Elliot Callender25 Mar 2026 22:07 UTC
12 points
0 comments3 min readLW link

PauseAI Capi­tol Day of Action

maia25 Mar 2026 21:21 UTC
9 points
0 comments1 min readLW link

A Toy En­vi­ron­ment For Ex­plor­ing Rea­son­ing About Reward

25 Mar 2026 20:29 UTC
60 points
10 comments3 min readLW link

Im­mor­tal­ity: A Beg­giner’s Guide (Part 3)

MarkelKori25 Mar 2026 19:10 UTC
4 points
0 comments4 min readLW link

Find­ing X-Risks and S-Risks by Gra­di­ent Descent

dspeyer25 Mar 2026 17:58 UTC
5 points
2 comments2 min readLW link

The Scary Bridge

moridinamael25 Mar 2026 17:11 UTC
16 points
0 comments2 min readLW link

How to do illu­sion­ist meditation

Jack Thompson25 Mar 2026 17:10 UTC
9 points
0 comments11 min readLW link

Can Agents Fool Each Other? Find­ings from the AI Village

Shoshannah Tekofsky25 Mar 2026 17:05 UTC
71 points
7 comments3 min readLW link
(theaidigest.org)

I lost my faith in in­tro­spec­tion—and you can too!

Jack Thompson25 Mar 2026 17:00 UTC
9 points
8 comments4 min readLW link

Don’t Write Off Hu­man La­bor, Yet

burnssa25 Mar 2026 16:37 UTC
−3 points
0 comments8 min readLW link

ECL-pilled mod­els write con­sti­tu­tions for ASI

kanad25 Mar 2026 16:05 UTC
16 points
0 comments16 min readLW link

Uncer­tain Up­dates: March 2026

Gordon Seidoh Worley25 Mar 2026 16:00 UTC
10 points
0 comments1 min readLW link
(www.uncertainupdates.com)

Claude Code, Cowork and Codex #6: Claude Code Auto Use and Full Cowork Com­puter Use

Zvi25 Mar 2026 15:10 UTC
42 points
0 comments17 min readLW link
(thezvi.wordpress.com)

How to do cost-effec­tive­ness anal­y­sis for elections

Zach Stein-Perlman25 Mar 2026 15:00 UTC
28 points
5 comments3 min readLW link

$1 billion is not enough; OpenAI Foun­da­tion must start spend­ing tens of billions each year

Davidmanheim25 Mar 2026 13:04 UTC
49 points
11 comments1 min readLW link

Every ACX/​LW House Party

Ravenstales25 Mar 2026 7:16 UTC
18 points
1 comment10 min readLW link

My Cog­ni­tive Ar­chi­tec­ture: A Self-Ob­ser­va­tional Map

Naj Ami-Nave25 Mar 2026 2:23 UTC
1 point
1 comment8 min readLW link

A Span­ish-Speak­ing Robot in my Pocket

jefftk25 Mar 2026 2:20 UTC
20 points
0 comments3 min readLW link
(www.jefftk.com)

Is Gem­ini 3 Schem­ing in the Wild?

25 Mar 2026 1:12 UTC
81 points
5 comments17 min readLW link

AI 2027 ver­sus World War 2027

Mitchell_Porter24 Mar 2026 23:57 UTC
9 points
6 comments2 min readLW link

Book Re­view: Open Socrates (Part 1)

Zvi24 Mar 2026 22:21 UTC
32 points
4 comments116 min readLW link
(thezvi.wordpress.com)

Book Re­view: Open Socrates (Part 2)

Zvi24 Mar 2026 22:20 UTC
22 points
2 comments81 min readLW link
(thezvi.wordpress.com)

Agents Can Get Stuck in Self-dis­trust­ing Equilibria

Ashe Vazquez Nuñez24 Mar 2026 22:05 UTC
33 points
2 comments12 min readLW link

La­tent In­tro­spec­tion (and other open-source in­tro­spec­tion pa­pers)

24 Mar 2026 21:23 UTC
98 points
3 comments9 min readLW link
(arxiv.org)

An In­for­mal Defi­ni­tion of Goals for Embed­ded Agents

Ashe Vazquez Nuñez24 Mar 2026 18:36 UTC
14 points
0 comments1 min readLW link

My cost-effec­tive­ness unit

Zach Stein-Perlman24 Mar 2026 15:30 UTC
65 points
5 comments4 min readLW link

AI Safety Newslet­ter #70: Au­to­mated War­fare and AI Layoffs

24 Mar 2026 15:30 UTC
8 points
0 comments4 min readLW link
(newsletter.safe.ai)

Mon­day AI Radar #18

Against Moloch24 Mar 2026 15:15 UTC
7 points
4 comments8 min readLW link
(againstmoloch.com)

The Fourth World

Linch24 Mar 2026 13:43 UTC
27 points
16 comments6 min readLW link

Safe Re­cur­sive Self-Im­prove­ment with Ver­ified Compilers

Adam Chlipala24 Mar 2026 13:35 UTC
15 points
0 comments11 min readLW link

Com­par­ing Across Pos­si­ble Worlds

unruly abstractions24 Mar 2026 10:09 UTC
7 points
4 comments5 min readLW link

Com­ing of Age: Chap­ters 1 and 2

Ihor Kendiukhov24 Mar 2026 9:15 UTC
13 points
0 comments1 min readLW link

The AIXI per­spec­tive on AI Safety

Cole Wyeth24 Mar 2026 3:24 UTC
80 points
4 comments6 min readLW link

Con­tra Dances Should Avoid Saturdays

jefftk24 Mar 2026 2:30 UTC
11 points
0 comments1 min readLW link
(www.jefftk.com)

Malmö AI Safety meetup

Vadym Sulzhenko (Vaigotaku)24 Mar 2026 2:29 UTC
1 point
0 comments1 min readLW link

In­for­ma­tion Overdose

Tridiv Sharma24 Mar 2026 2:26 UTC
−2 points
0 comments3 min readLW link

We can­not safely au­to­mate value al­ign­ment eval­u­a­tion and re­search with­out think­ing about del­e­ga­tion and discretion

Mflena24 Mar 2026 2:25 UTC
8 points
0 comments10 min readLW link

Every Ma­jor LLM is a 1-Box Smok­ing Thirder

Olivia Scharfman24 Mar 2026 2:18 UTC
14 points
4 comments10 min readLW link

A ToM-In­spired Agenda for AI Safety Research

Andrés Cotton24 Mar 2026 2:13 UTC
7 points
1 comment5 min readLW link

Pok­ing and Edit­ing the Circuits

unruly abstractions24 Mar 2026 1:15 UTC
2 points
0 comments8 min readLW link

Adopt a de­bug­ger’s mind­set to solve your re­cur­ring life problems

Declan Molony24 Mar 2026 0:22 UTC
24 points
6 comments8 min readLW link

Hol­ly­wood ACX Meetup—April 2026

Timothy M.23 Mar 2026 23:55 UTC
1 point
0 comments1 min readLW link

Ex­per­i­men­tal Ev­i­dence for Si­mu­la­tor The­ory— Part 2: The Scalers Strike Back

RogerDearnaley23 Mar 2026 22:37 UTC
21 points
0 comments34 min readLW link

Ex­per­i­men­tal Ev­i­dence for Si­mu­la­tor The­ory— Part 1: Emer­gent Misal­ign­ment and Weird Generalizations

RogerDearnaley23 Mar 2026 22:37 UTC
25 points
0 comments53 min readLW link

How Far Can Ob­ser­va­tion Take Us?

unruly abstractions23 Mar 2026 21:56 UTC
12 points
0 comments9 min readLW link

Vibe­coders can’t build for longevity

dominicq23 Mar 2026 19:36 UTC
13 points
9 comments4 min readLW link

Ablat­ing Split Per­son­al­ity Training

OscarGilg23 Mar 2026 17:45 UTC
55 points
1 comment5 min readLW link

Mea­sur­ing and im­prov­ing cod­ing au­dit re­al­ism with de­ploy­ment resources

23 Mar 2026 17:20 UTC
43 points
1 comment10 min readLW link
(alignment.anthropic.com)

AI char­ac­ter is a big deal

23 Mar 2026 16:36 UTC
34 points
33 comments12 min readLW link
(www.forethought.org)

Some things I no­ticed while LARPing as a grantmaker

Zach Stein-Perlman23 Mar 2026 15:00 UTC
163 points
11 comments7 min readLW link