How Go Play­ers Disem­power Them­selves to AI

Ashe Vazquez Nuñez1 May 2026 23:24 UTC
746 points
79 comments8 min readLW link

The Owned Ones

Eliezer Yudkowsky12 May 2026 17:56 UTC
399 points
53 comments6 min readLW link

Women should be able to open things

KatjaGrace21 May 2026 3:50 UTC
379 points
142 comments2 min readLW link
(worldspiritsockpuppet.com)

Ir­re­triev­abil­ity; or, Mur­phy’s Curse of Oneshot­ness upon ASI

Eliezer Yudkowsky4 May 2026 22:11 UTC
370 points
132 comments22 min readLW link

Mnemonic por­traits for 19,023 hu­man genes

Brinedew28 May 2026 22:16 UTC
363 points
29 comments15 min readLW link

It’s nice of you to worry about me, but I re­ally do have a life

Viliam4 May 2026 21:14 UTC
345 points
66 comments4 min readLW link

Trees are mostly made of air and a gen­er­al­iz­able les­son for AI safety

Zephaniah Roe29 May 2026 4:08 UTC
324 points
62 comments4 min readLW link

Models find­ing soft­ware vuln­er­a­bil­ities is not the pri­mary source of cy­ber­se­cu­rity risk

Dean Valentine14 May 2026 3:39 UTC
316 points
25 comments2 min readLW link

Bad Prob­lems Don’t Stop Be­ing Bad Be­cause Some­body’s Wrong About Fault Analysis

Linch9 May 2026 1:30 UTC
268 points
75 comments3 min readLW link

x-risk-themed

kave6 May 2026 15:16 UTC
251 points
26 comments3 min readLW link
(kaverennedy.substack.com)

A rel­a­tively brief ex­pla­na­tion of Boltz­mann Brains

Eliezer Yudkowsky16 May 2026 21:19 UTC
216 points
162 comments4 min readLW link

MATS 9 Ret­ro­spec­tive & Advice

beyarkay (Boyd Kane)15 May 2026 12:30 UTC
212 points
13 comments18 min readLW link
(boydkane.com)

Nat­u­ral Lan­guage Au­toen­coders Pro­duce Un­su­per­vised Ex­pla­na­tions of LLM Activations

7 May 2026 20:21 UTC
209 points
35 comments8 min readLW link

Em­pow­er­ment, cor­rigi­bil­ity, etc. are sim­ple ab­strac­tions (of a messed-up on­tol­ogy)

Steven Byrnes11 May 2026 17:48 UTC
193 points
74 comments16 min readLW link

Who Got Breasts First and How We Got Them

rba11 May 2026 13:11 UTC
165 points
59 comments10 min readLW link

A Year Late, Claude Fi­nally Beats Poké­mon

Julian Bradshaw16 May 2026 7:05 UTC
164 points
14 comments9 min readLW link

[Linkpost] In­ter­pret­ing Lan­guage Model Parameters

5 May 2026 17:37 UTC
164 points
2 comments2 min readLW link
(www.goodfire.ai)

Cog­ni­tive Se­cu­rity as an AI Safety Cause Area

jsteinhardt25 May 2026 18:30 UTC
162 points
20 comments2 min readLW link

Dairy cows make their mis­ery ex­pen­sive (but their calves can’t)

Elizabeth3 May 2026 19:20 UTC
161 points
4 comments6 min readLW link
(acesounderglass.com)

The Iliad In­ten­sive Course Ma­te­ri­als (April 2026)

11 May 2026 18:55 UTC
160 points
4 comments13 min readLW link
(docs.google.com)

Au­to­mated Align­ment is Harder Than You Think

14 May 2026 22:01 UTC
145 points
8 comments3 min readLW link
(arxiv.org)

The Dar­wi­nian Honey­moon—Why I am not as im­pressed by hu­man progress as I used to be

Elias Schmied10 May 2026 15:55 UTC
144 points
23 comments4 min readLW link

the­ory up­lift differ­en­tially benefits safety & is underleveraged

yudhister20 May 2026 21:43 UTC
130 points
14 comments1 min readLW link

Risk from fit­ness-seek­ing AIs: mechanisms and mitigations

Alex Mallen1 May 2026 17:42 UTC
130 points
0 comments32 min readLW link

Syn­thetic Per­sona Pre­train­ing: Align­ment from To­ken Zero

20 May 2026 14:16 UTC
130 points
27 comments17 min readLW link
(modelraising.ai)

You Are Not Im­mune To Mode Collapse

J Bostock2 May 2026 19:57 UTC
128 points
18 comments4 min readLW link
(jbostock.substack.com)

Con­tra Went­worth on Phys­i­cal At­trac­tive­ness for Men

Gretta Duleba26 May 2026 23:20 UTC
127 points
28 comments8 min readLW link

The AI In­dus­trial Ex­plo­sion — Part 1: Max­i­mum growth rates with cur­rent pro­duc­tion methods

djbinder4 May 2026 15:32 UTC
126 points
14 comments12 min readLW link
(defensesindepth.bio)

Tak­ing woo se­ri­ously but not literally

Kaj_Sotala4 May 2026 13:36 UTC
126 points
27 comments23 min readLW link
(kajsotala.substack.com)

Con­ver­gent Ab­strac­tion Hypothesis

Jan_Kulveit15 May 2026 0:04 UTC
125 points
20 comments6 min readLW link

Donat­ing 80% While It Still Counts

jefftk26 May 2026 1:30 UTC
123 points
9 comments6 min readLW link
(www.jefftk.com)

Ne­ga­tion Ne­glect: When mod­els fail to learn nega­tions in training

18 May 2026 18:37 UTC
121 points
37 comments8 min readLW link

Claude, Author of the Humanitas

Linch26 May 2026 16:05 UTC
119 points
42 comments16 min readLW link

In­crim­i­nat­ing mis­al­igned AI mod­els via distillation

15 May 2026 21:43 UTC
119 points
12 comments5 min readLW link

Op­ti­mi­sa­tion: Selec­tive ver­sus Predictive

Raymond Douglas12 May 2026 14:03 UTC
118 points
15 comments3 min readLW link

Vot­ers are sur­pris­ingly open to talk­ing about AI risk

less_raichu13 May 2026 14:08 UTC
117 points
11 comments3 min readLW link

Im­pli­ca­tions Of Pre­dict­ing The Next Token

jdp19 May 2026 22:17 UTC
115 points
6 comments31 min readLW link
(minihf.com)

Many in­di­vi­d­ual CEVs are prob­a­bly quite bad

Viliam6 May 2026 20:18 UTC
111 points
32 comments3 min readLW link

Try, even if they have you cold

WalterL7 May 2026 17:19 UTC
105 points
14 comments2 min readLW link

In­ter­na­tional Law Can­not Prevent Ex­tinc­tion Either

Sausage Vector Machine9 May 2026 22:34 UTC
104 points
16 comments5 min readLW link

Prac­ti­cal Learn­ings from Syn­thetic Doc­u­ment Finetuning

26 May 2026 19:22 UTC
104 points
7 comments8 min readLW link

Prob­a­bil­ities are not the right concept

David Matolcsi23 May 2026 16:10 UTC
94 points
33 comments15 min readLW link

Will we re­ally put data cen­ters in space?

22 May 2026 23:51 UTC
94 points
28 comments5 min readLW link
(www.forethought.org)

Don’t be too Clever to Take Ob­vi­ous Ad­vice

Hide15 May 2026 3:01 UTC
93 points
26 comments2 min readLW link
(hidefromit.substack.com)

Your rights when fly­ing to Europe

Yair Halberstadt5 May 2026 19:17 UTC
92 points
15 comments5 min readLW link

Bring­ing More Ex­per­tise to Bear on Alignment

8 May 2026 10:29 UTC
92 points
1 comment8 min readLW link

Tax­ing Small Cars To Im­prove MPG

jefftk24 May 2026 21:50 UTC
92 points
11 comments2 min readLW link
(www.jefftk.com)

Claude is Now Align­ment-Pretrained

RogerDearnaley13 May 2026 23:19 UTC
90 points
9 comments1 min readLW link
(www.anthropic.com)

What am I, if not an AI?

makiba21 May 2026 13:14 UTC
85 points
15 comments7 min readLW link

An­nounc­ing Geodesic Research

27 May 2026 16:40 UTC
85 points
2 comments5 min readLW link