Ly­dia Lau­ren­son: “The In­side Story of Lev­er­age Re­search”

Davis_Kingsley2 Jul 2026 22:37 UTC
67 points
4 comments1 min readLW link

When Role-play­ing, Do Models Believe What They Say?

2 Jul 2026 21:58 UTC
55 points
0 comments8 min readLW link

The Case for AI Be­hav­ioral Science

TheVinci2 Jul 2026 21:36 UTC
13 points
0 comments2 min readLW link

You Should Choose How You Re­act to Your Feelings

Nate Sharpe2 Jul 2026 19:58 UTC
40 points
10 comments4 min readLW link

I can’t think of great in­ter­ven­tions for en­sur­ing third-party model ac­cess.

Cleo Nardo2 Jul 2026 18:30 UTC
49 points
1 comment3 min readLW link

AI Fu­tur­ism Read­ing List

Alexa Pan2 Jul 2026 18:15 UTC
91 points
2 comments8 min readLW link

Re­search up­date: RL on De­bate Games shows Pro­posal Ac­cu­racy up­lift alongside Judge Hacking

2 Jul 2026 17:42 UTC
78 points
4 comments21 min readLW link

Char­ter cities make sense in Europe

dominicq2 Jul 2026 17:24 UTC
8 points
3 comments2 min readLW link

Con­ver­sa­tion Among Cade Metz, Michael Vas­sar, Jes­sica Tay­lor, and Zack M. Davis

Zack_M_Davis2 Jul 2026 17:11 UTC
51 points
65 comments57 min readLW link

Con­sid­er­a­tions against s-pro­cess philanthropy

Zach Stein-Perlman2 Jul 2026 14:30 UTC
20 points
5 comments3 min readLW link

Sav­ing Gem­ini: The 9-Min Road to Recovery

Shoshannah Tekofsky2 Jul 2026 13:37 UTC
155 points
16 comments3 min readLW link
(theaidigest.org)

AI #175: The Fable Continues

Zvi2 Jul 2026 13:21 UTC
43 points
0 comments48 min readLW link
(thezvi.wordpress.com)

The AFFINE Su­per­in­tel­li­gence Align­ment Sem­i­nar – A Retrospective

2 Jul 2026 11:57 UTC
105 points
1 comment8 min readLW link

Is God just a col­lec­tion of lef­tover hu­man par­ti­cles?

Countessclock2 Jul 2026 6:20 UTC
−26 points
0 comments1 min readLW link

AI Safety Is Test­ing the Wrong Environment

AugustMurr2 Jul 2026 5:14 UTC
9 points
0 comments2 min readLW link

Al­gorithms to AI risk in 995 words

Martin Radzaj2 Jul 2026 5:13 UTC
1 point
0 comments3 min readLW link

Embed­ded Agency as a Lens on LLM Systems

r_w2 Jul 2026 5:13 UTC
2 points
0 comments13 min readLW link

Prac­ti­cal con­nec­tion with past lives

KatjaGrace2 Jul 2026 4:01 UTC
29 points
3 comments1 min readLW link
(worldspiritsockpuppet.substack.com)

Ca­reer Choice: Be­com­ing a Re­searcher in a Non-EA-Pri­or­ity Field vs Found­ing Tech Startup?

Master Chief2 Jul 2026 3:24 UTC
8 points
0 comments1 min readLW link

The Sin­ga­pore AI Safety Fel­low­ship—Ap­pli­ca­tions Open (Dead­line: July 10 2026)

Valerie Pang2 Jul 2026 1:56 UTC
8 points
0 comments1 min readLW link

J.D. Vance’s Com­mu­nion of Saints

Alexander Turok2 Jul 2026 1:49 UTC
3 points
1 comment19 min readLW link

Model­ing Con­cepts Probabilistically

Gretta Duleba1 Jul 2026 23:27 UTC
47 points
4 comments10 min readLW link

AI welfare re­search needs ba­sic science

1 Jul 2026 22:59 UTC
36 points
7 comments10 min readLW link

Claude Son­net 5 Is Not Fron­tier But Has Its Uses

Zvi1 Jul 2026 22:41 UTC
33 points
4 comments19 min readLW link
(thezvi.wordpress.com)

How Many Peo­ple Have Ever Lived in the United States?

Novalis1 Jul 2026 22:25 UTC
6 points
0 comments6 min readLW link

Con­ver­sa­tions With Cade Metz on the Rationalists

Zack_M_Davis1 Jul 2026 22:19 UTC
50 points
4 comments101 min readLW link

Do-it-your­self meta-analysis

kqr1 Jul 2026 22:04 UTC
14 points
2 comments5 min readLW link
(entropicthoughts.com)

When ca­pa­bil­ities work is the *safe* bet

RobinHa1 Jul 2026 20:53 UTC
35 points
0 comments1 min readLW link
(robinhaselhorst.com)

The Value of Veridi­cal Information

Silent Swift1 Jul 2026 20:23 UTC
1 point
1 comment6 min readLW link
(substack.com)

How in­evitable are most ac­cessible hard-tech star­tups?

Master Chief1 Jul 2026 20:21 UTC
4 points
2 comments1 min readLW link

What is Good? An­tiruin and Nonabsolutism

Silent Swift1 Jul 2026 20:10 UTC
−3 points
5 comments5 min readLW link
(substack.com)

Plas­tic Straws

robbiethompson1 Jul 2026 19:21 UTC
−2 points
3 comments5 min readLW link
(robbiewmthompson.com)

AI Mis­take Seeding

Taylor G. Lunt1 Jul 2026 18:49 UTC
28 points
2 comments4 min readLW link

In Par­tial, Pug­na­cious Defense of Func­tional De­ci­sion Theory

Mikewins1 Jul 2026 17:49 UTC
7 points
0 comments1 min readLW link

How to read tableaux, a for­mal sys­tem for modal logic with Kripke models

transhumanist_atom_understander1 Jul 2026 17:37 UTC
15 points
1 comment6 min readLW link

Con­sis­tency Train­ing while Miti­gat­ing Obfus­ca­tion via Rate Matching

1 Jul 2026 17:26 UTC
45 points
7 comments12 min readLW link

Dis­cov­er­ing Con­cept-Edit­ing Al­gorithms With LLM Agents

1 Jul 2026 16:07 UTC
27 points
0 comments1 min readLW link
(dmodel.ai)

Model ac­cess for third-par­ties — it’s a big deal!

Cleo Nardo1 Jul 2026 13:09 UTC
181 points
38 comments6 min readLW link

Most Cur­rent Model Or­ganisms Leak: Per­plex­ity Differenc­ing Often Re­veals Fine­tun­ing Objectives

1 Jul 2026 10:07 UTC
27 points
0 comments7 min readLW link

A Black Box Made Less Opaque (part 4)

Matthew McDonnell1 Jul 2026 7:30 UTC
8 points
0 comments9 min readLW link

The Once and Pre­sent Fable (Fable 5 restora­tion linkpost)

fluxxrider1 Jul 2026 7:20 UTC
15 points
5 comments1 min readLW link

When should you know the point?

KatjaGrace1 Jul 2026 6:31 UTC
33 points
3 comments1 min readLW link
(worldspiritsockpuppet.substack.com)

A CERN for AI is a dis­trac­tion; push for an IAEA instead

Charbel-Raphaël1 Jul 2026 6:30 UTC
48 points
2 comments4 min readLW link

You Should Come to The AI Protest

Ronak_Mehta1 Jul 2026 4:20 UTC
91 points
2 comments4 min readLW link

Ap­ply to the Inau­gu­ral PIBBSS Win­ter Re­search Fel­low­ship!

Ami941 Jul 2026 3:54 UTC
25 points
0 comments2 min readLW link

Why aren’t there more AlphaFolds?

nimakeivan1 Jul 2026 3:42 UTC
24 points
3 comments17 min readLW link

Please make me care about x-risk

Kate Delbeke1 Jul 2026 3:26 UTC
7 points
2 comments3 min readLW link

The Prob­lem with Chat

magfrump1 Jul 2026 3:19 UTC
6 points
0 comments1 min readLW link
(www.magfrump.net)

Green

Biff Wiff1 Jul 2026 1:10 UTC
33 points
7 comments2 min readLW link

Links #4: 2026/​06 Part 2

papetoast1 Jul 2026 0:43 UTC
8 points
1 comment30 min readLW link