Hol­ly­wood ACX Meetup—April 2026

Timothy M.23 Mar 2026 23:55 UTC
1 point
0 comments1 min readLW link

Ex­per­i­men­tal Ev­i­dence for Si­mu­la­tor The­ory— Part 2: The Scalers Strike Back

RogerDearnaley23 Mar 2026 22:37 UTC
21 points
0 comments34 min readLW link

Ex­per­i­men­tal Ev­i­dence for Si­mu­la­tor The­ory— Part 1: Emer­gent Misal­ign­ment and Weird Generalizations

RogerDearnaley23 Mar 2026 22:37 UTC
25 points
0 comments53 min readLW link

How Far Can Ob­ser­va­tion Take Us?

unruly abstractions23 Mar 2026 21:56 UTC
12 points
0 comments9 min readLW link

Vibe­coders can’t build for longevity

dominicq23 Mar 2026 19:36 UTC
13 points
9 comments4 min readLW link

Ablat­ing Split Per­son­al­ity Training

OscarGilg23 Mar 2026 17:45 UTC
55 points
1 comment5 min readLW link

Mea­sur­ing and im­prov­ing cod­ing au­dit re­al­ism with de­ploy­ment resources

23 Mar 2026 17:20 UTC
43 points
1 comment10 min readLW link
(alignment.anthropic.com)

AI char­ac­ter is a big deal

23 Mar 2026 16:36 UTC
34 points
33 comments12 min readLW link
(www.forethought.org)

Some things I no­ticed while LARPing as a grantmaker

Zach Stein-Perlman23 Mar 2026 15:00 UTC
163 points
11 comments7 min readLW link

Rep­re­sen­ta­tive Futarchy

goldfine23 Mar 2026 14:39 UTC
2 points
0 comments2 min readLW link
(itsnotgambling.substack.com)

Which types of AI al­ign­ment re­search are most likely to be good for all sen­tient be­ings?

MichaelDickens23 Mar 2026 13:38 UTC
11 points
1 comment6 min readLW link

Kelly Cri­te­rion is for Cowards

X4vier23 Mar 2026 1:56 UTC
16 points
28 comments3 min readLW link

Set the Line Be­fore It’s Crossed

nomagicpill23 Mar 2026 1:25 UTC
12 points
0 comments5 min readLW link
(nomagicpill.substack.com)

When Align­ment Be­comes an At­tack Sur­face: Prompt In­jec­tion in Co­op­er­a­tive Multi-Agent Systems

Cornelis Dirk Haupt23 Mar 2026 0:22 UTC
9 points
0 comments4 min readLW link

At­tend the 2026 Re­pro­duc­tive Fron­tiers Sum­mit, June 16–18, Berkeley

22 Mar 2026 21:15 UTC
97 points
1 comment5 min readLW link

You’re ab­solutely right, Se­na­tor. I was be­ing naive about the poli­ti­cal re­al­ity.

Chris Datcu22 Mar 2026 20:42 UTC
4 points
5 comments2 min readLW link

Let’s Rea­son About (Your) Job Se­cu­rity!

Gergely Máté22 Mar 2026 20:13 UTC
8 points
0 comments4 min readLW link

Madi­son ACX Mee­tups Every­where

edalva22 Mar 2026 17:22 UTC
1 point
0 comments1 min readLW link

Is fever a symp­tom of glycine defi­ciency?

Benquo22 Mar 2026 14:44 UTC
218 points
86 comments6 min readLW link
(benjaminrosshoffman.com)

My Most Costly Delusion

Ihor Kendiukhov22 Mar 2026 12:21 UTC
161 points
14 comments3 min readLW link

Notic­ing a Teacher’s Pass­word Pattern

Dentosal22 Mar 2026 9:10 UTC
26 points
12 comments2 min readLW link

The Toy Story Saga is not yet finished

Raemon22 Mar 2026 3:53 UTC
95 points
9 comments8 min readLW link

Key to Life No. 9: Access

MarkelKori21 Mar 2026 21:53 UTC
11 points
0 comments3 min readLW link

My Ham­mer­time Fi­nal Exam

evjeny21 Mar 2026 20:40 UTC
11 points
0 comments2 min readLW link

Un­der­stand­ing when and why agents scheme

21 Mar 2026 20:33 UTC
50 points
2 comments4 min readLW link

Build­ing a Web App Us­ing an AI-As­sisted Workflow

Thomas Castriensis21 Mar 2026 19:21 UTC
1 point
0 comments1 min readLW link

China Derange­ment Syndrome

Arjun Panickssery21 Mar 2026 19:19 UTC
124 points
70 comments4 min readLW link
(arjunpanickssery.substack.com)

China de­clares AGI de­vel­op­ment to be a part of 5-year plan

Darmani21 Mar 2026 17:21 UTC
32 points
4 comments1 min readLW link

Utrecht Meetup #2, Mak­ing Beliefs Pay Rent

aad21 Mar 2026 16:12 UTC
4 points
1 comment1 min readLW link

Ground­ing Cod­ing Agents via Dixit

qbolec21 Mar 2026 11:01 UTC
15 points
0 comments10 min readLW link

The Hot Mess Paper Con­flates Three Distinct Failure Modes

laudiacay21 Mar 2026 2:57 UTC
28 points
3 comments6 min readLW link

The Fu­ture of Align­ing Deep Learn­ing sys­tems will prob­a­bly look like “train­ing on in­terp”

williawa20 Mar 2026 23:06 UTC
29 points
7 comments4 min readLW link

An agent au­tonomously builds a 1.5 GHz Linux-ca­pa­ble RISC-V CPU

sanxiyn20 Mar 2026 23:03 UTC
19 points
2 comments2 min readLW link
(arxiv.org)

Un­trusted mon­i­tor­ing: ex­tra bits

Morgan S20 Mar 2026 21:32 UTC
26 points
0 comments15 min readLW link

Find­ing fea­tures in Trans­form­ers: Con­trastive di­rec­tions elicit stronger low-level per­tur­ba­tion re­sponses than baselines

20 Mar 2026 21:09 UTC
39 points
2 comments6 min readLW link

ARENA 7.0 Im­pact Report

20 Mar 2026 17:09 UTC
13 points
0 comments21 min readLW link

The Fed­eral AI Policy Frame­work: An Im­prove­ment, But My Offer Is (Still Al­most) Nothing

Zvi20 Mar 2026 16:51 UTC
33 points
0 comments8 min readLW link
(thezvi.wordpress.com)

Con­fu­sion around the term re­ward hacking

ariana_azarbal20 Mar 2026 16:13 UTC
67 points
6 comments5 min readLW link

The Distaff Texts

Tomás B.20 Mar 2026 15:05 UTC
105 points
6 comments14 min readLW link

It’s a Good Thing to Re­spond to In­ter­net Trolls

Bowl of Cereal20 Mar 2026 14:22 UTC
−10 points
4 comments2 min readLW link

Un­trusted Mon­i­tor­ing is De­fault; Trusted Mon­i­tor­ing is not

J Bostock20 Mar 2026 14:10 UTC
30 points
0 comments4 min readLW link

Against Mes­si­anic AI: Why Op­ti­miz­ing the En­vi­ron­ment Doesn’t Op­ti­mize the Agent

Nathan Heath20 Mar 2026 12:40 UTC
1 point
0 comments3 min readLW link

2nd (Unoffi­cial) ACX Weekend

Fernand020 Mar 2026 12:13 UTC
1 point
0 comments1 min readLW link

Why I am not buy­ing IPv4 ad­dresses as an investment

samuelshadrach20 Mar 2026 9:02 UTC
4 points
2 comments5 min readLW link
(samuelshadrach.com)

Hun­dred ways a su­per­in­tel­li­gence could kill you (non-se­ri­ous ex­er­cise)

samuelshadrach20 Mar 2026 8:58 UTC
3 points
1 comment6 min readLW link
(samuelshadrach.com)

In­ter­net anonymity with­out Tor

samuelshadrach20 Mar 2026 8:52 UTC
1 point
0 comments3 min readLW link
(samuelshadrach.com)

No, You Don’t Need Self-Lo­cat­ing Ev­i­dence.

Ape in the coat20 Mar 2026 5:38 UTC
8 points
4 comments5 min readLW link
(substack.com)

The Low Hang­ing Fruit of AI Self Improvement

HunterJay20 Mar 2026 4:09 UTC
1 point
0 comments5 min readLW link

Nul­lius in Verba: 3rd party ev­i­dence for Nec­tome’s Brain Preservation

Aurelia20 Mar 2026 3:19 UTC
179 points
18 comments12 min readLW link

Does He­brew Have Verbs?

Benquo20 Mar 2026 3:04 UTC
37 points
9 comments6 min readLW link
(benjaminrosshoffman.com)