RSS

Should safety re­searchers quit fron­tier labs?

Ryan Kidd5 Sep 2026 1:17 UTC
56 points
9 comments3 min readLW link

AI Librar­i­ans Lower the Bar for Shar­ing Your Writing

utilistrutil4 Sep 2026 22:54 UTC
9 points
0 comments3 min readLW link

Does progress in AI safety re­quire progress in AI ca­pa­bil­ities?

Stephen McAleese4 Sep 2026 21:38 UTC
11 points
0 comments15 min readLW link

Let’s talk about the AI co­or­di­na­tion problem

KatjaGrace4 Sep 2026 20:16 UTC
47 points
5 comments2 min readLW link

Prob­ing for Cal­ibrated Rare Action

Zach Allen4 Sep 2026 20:01 UTC
2 points
0 comments3 min readLW link
(github.com)

Eat Me. Drink Me. Copy, Paste, and Run Me.

derelict54324 Sep 2026 19:31 UTC
−1 points
7 comments3 min readLW link

F***ing Pul­leys, How Do They Work?

Liron4 Sep 2026 16:57 UTC
19 points
11 comments3 min readLW link
(lironshapira.substack.com)

Safe(r) Self-Driv­ing Labs #1

Ana Leonescu4 Sep 2026 16:20 UTC
8 points
0 comments11 min readLW link

Train­ing Models to Pre­dict and Ex­plain Their In-the-Wild Behavior

4 Sep 2026 16:16 UTC
32 points
2 comments8 min readLW link

Al­most no­body is funded to figure out what work would solve alignment

Seth Herd4 Sep 2026 15:57 UTC
54 points
17 comments4 min readLW link

Dis­cov­ery Of A New OpenAI Agent Mes­sage Board

Capybasilisk4 Sep 2026 14:46 UTC
208 points
29 comments1 min readLW link
(collusion.wiki)

Meta—Physics I: Why don’t we live in the Game of Life?

interstice4 Sep 2026 12:16 UTC
9 points
1 comment6 min readLW link
(thermontology.com)

AI risk and the ra­tio­nal voter

djbinder4 Sep 2026 11:38 UTC
46 points
1 comment3 min readLW link
(defensesindepth.bio)

Ask­ing agents to make money to survive

invertedpassion4 Sep 2026 7:44 UTC
11 points
1 comment10 min readLW link

So­cietal im­pacts re­search has no timeline

emiliob4 Sep 2026 2:53 UTC
10 points
0 comments1 min readLW link

Higher ed­u­ca­tion as class commitment

Richard_Ngo4 Sep 2026 2:00 UTC
50 points
4 comments16 min readLW link
(www.mindthefuture.info)

Do gen­eral-pur­pose robots mean­ingfully in­crease ASI takeover risk?

Master Chief4 Sep 2026 0:53 UTC
2 points
2 comments1 min readLW link

Ab­strac­tion Equivocation

WillPetillo3 Sep 2026 23:28 UTC
12 points
0 comments4 min readLW link

A Ther­a­pist for Peo­ple Who Think the World Might End: An In­ter­view with Daystar Eld (Da­mon Sasi)

JohnGreer3 Sep 2026 22:14 UTC
6 points
0 comments49 min readLW link
(youtu.be)

Cat-Bel­ling Problems

Eliezer Yudkowsky3 Sep 2026 21:20 UTC
193 points
55 comments21 min readLW link