Get­ting an Oura ring im­proved my sleep and exercise

Daniel Tan5 Jul 2026 23:26 UTC
11 points
0 comments2 min readLW link

Fealty to Fidelity 👰

chaosmage5 Jul 2026 23:25 UTC
19 points
0 comments3 min readLW link

Reflec­tions on The Scout Mindset

James Brobin5 Jul 2026 20:18 UTC
12 points
0 comments3 min readLW link

When Gemma Thinks About Re­sources—it Fails: a Be­hav­ioral Experiment

TheVinci5 Jul 2026 19:23 UTC
8 points
2 comments2 min readLW link
(tarantulabs.com)

Claude’s mal­i­cious com­pli­ance and nor­mal­iza­tion of deviance

Steff5 Jul 2026 17:18 UTC
27 points
23 comments8 min readLW link

Book Re­view: The God Test

PeterMcCluskey5 Jul 2026 16:11 UTC
25 points
3 comments4 min readLW link

We need 3rd party Train­ing-Run Evaluations

Alex Meinke5 Jul 2026 15:55 UTC
176 points
2 comments11 min readLW link

Harry Pot­ter and the Rules of Quidditch

Tomás B.5 Jul 2026 14:32 UTC
161 points
8 comments3 min readLW link

A Nor­mal Ar­gu­ment for AI Risk

Silent Swift5 Jul 2026 9:32 UTC
16 points
3 comments8 min readLW link
(silentswift.substack.com)

The case for the fleshman

dr_s5 Jul 2026 9:28 UTC
21 points
20 comments4 min readLW link

Ree­val­u­at­ing AI-2027: timelines, take­off, al­ign­ment and China

StanislavKrym5 Jul 2026 4:00 UTC
16 points
6 comments5 min readLW link

Suc­cess Per Tokens

michaelwaves5 Jul 2026 2:25 UTC
8 points
0 comments3 min readLW link

Re­sults of a small ZBiotics RCT

Nikola Jurkovic5 Jul 2026 2:14 UTC
64 points
5 comments1 min readLW link

A case for LLMs as Self-predictors

Ashe Vazquez Nuñez5 Jul 2026 0:25 UTC
34 points
6 comments10 min readLW link

Ver­i­zon is About to Break our Watches

jefftk4 Jul 2026 17:50 UTC
18 points
3 comments2 min readLW link
(www.jefftk.com)

Defin­ing in­ter­pre­ta­tion, and es­tab­lish­ing a frame­work for it

Yaroven4 Jul 2026 16:31 UTC
6 points
0 comments5 min readLW link

Fluidity Fo­rum 2026

NormanPerlmutter4 Jul 2026 5:13 UTC
21 points
2 comments1 min readLW link

The Lace (short story)

Michael Soareverix4 Jul 2026 4:43 UTC
26 points
2 comments4 min readLW link

Ap­prox­i­mate Nat­u­ral La­tents Have Ex­act Prices

Haru4 Jul 2026 1:57 UTC
25 points
0 comments6 min readLW link

I think al­ign­ment work is more promis­ing than con­trol work

Alec Harris3 Jul 2026 23:40 UTC
102 points
15 comments8 min readLW link

On “gen­dertropes” in dath ilan

Eliezer Yudkowsky3 Jul 2026 22:20 UTC
78 points
1 comment3 min readLW link

Amer­i­can AI if the boom is a bub­ble: the Karp-Zitron scenario

Mitchell_Porter3 Jul 2026 21:46 UTC
12 points
1 comment2 min readLW link

(Don’t fear) the strangelet

djbinder3 Jul 2026 17:39 UTC
135 points
22 comments22 min readLW link
(defensesindepth.bio)

The Re­v­erse AI Box

James_Miller3 Jul 2026 16:08 UTC
9 points
2 comments6 min readLW link

An­nounc­ing the Safe Pareto Im­prove­ments (SPI) Fun­da­men­tals Program

Anthony DiGiovanni3 Jul 2026 15:55 UTC
51 points
1 comment3 min readLW link

Prag­matic FDT, and pre­dic­tors as game theory

Stuart_Armstrong3 Jul 2026 13:22 UTC
36 points
12 comments11 min readLW link

Fable #6: The Re­turn of the King

Zvi3 Jul 2026 13:22 UTC
51 points
1 comment13 min readLW link
(thezvi.wordpress.com)

June-July 2026 AI Se­cu­rity via For­mal Methods

Quinn3 Jul 2026 12:32 UTC
14 points
0 comments2 min readLW link
(newsletter.for-all.dev)

Schem­ing Evals Mislead in Both Directions

3 Jul 2026 11:49 UTC
22 points
0 comments10 min readLW link

Frag­ile Cor­rect­ness: Cases of rea­son­ing harm­ing performance

tobypullan3 Jul 2026 9:32 UTC
21 points
2 comments5 min readLW link

One axis and two fea­tures, how I solved the first puz­zle from BlueDot and how a clas­sifier hid coun­try on the food direction

IgorPereverzevDev3 Jul 2026 4:52 UTC
11 points
0 comments12 min readLW link

Ly­dia Lau­ren­son: “The In­side Story of Lev­er­age Re­search”

Davis_Kingsley2 Jul 2026 22:37 UTC
67 points
4 comments1 min readLW link

When Role-play­ing, Do Models Believe What They Say?

2 Jul 2026 21:58 UTC
55 points
0 comments8 min readLW link

The Case for AI Be­hav­ioral Science

TheVinci2 Jul 2026 21:36 UTC
13 points
0 comments2 min readLW link

You Should Choose How You Re­act to Your Feelings

Nate Sharpe2 Jul 2026 19:58 UTC
40 points
10 comments4 min readLW link

I can’t think of great in­ter­ven­tions for en­sur­ing third-party model ac­cess.

Cleo Nardo2 Jul 2026 18:30 UTC
49 points
1 comment3 min readLW link

AI Fu­tur­ism Read­ing List

Alexa Pan2 Jul 2026 18:15 UTC
92 points
2 comments8 min readLW link

Re­search up­date: RL on De­bate Games shows Pro­posal Ac­cu­racy up­lift alongside Judge Hacking

2 Jul 2026 17:42 UTC
78 points
4 comments21 min readLW link

Char­ter cities make sense in Europe

dominicq2 Jul 2026 17:24 UTC
9 points
3 comments2 min readLW link

Con­ver­sa­tion Among Cade Metz, Michael Vas­sar, Jes­sica Tay­lor, and Zack M. Davis

Zack_M_Davis2 Jul 2026 17:11 UTC
51 points
65 comments57 min readLW link

Con­sid­er­a­tions against s-pro­cess philanthropy

Zach Stein-Perlman2 Jul 2026 14:30 UTC
20 points
5 comments3 min readLW link

Sav­ing Gem­ini: The 9-Min Road to Recovery

Shoshannah Tekofsky2 Jul 2026 13:37 UTC
156 points
16 comments3 min readLW link
(theaidigest.org)

AI #175: The Fable Continues

Zvi2 Jul 2026 13:21 UTC
43 points
0 comments48 min readLW link
(thezvi.wordpress.com)

The AFFINE Su­per­in­tel­li­gence Align­ment Sem­i­nar – A Retrospective

2 Jul 2026 11:57 UTC
105 points
1 comment8 min readLW link

Is God just a col­lec­tion of lef­tover hu­man par­ti­cles?

Countessclock2 Jul 2026 6:20 UTC
−26 points
0 comments1 min readLW link

AI Safety Is Test­ing the Wrong Environment

AugustMurr2 Jul 2026 5:14 UTC
9 points
0 comments2 min readLW link

Al­gorithms to AI risk in 995 words

Martin Radzaj2 Jul 2026 5:13 UTC
1 point
0 comments3 min readLW link

Embed­ded Agency as a Lens on LLM Systems

r_w2 Jul 2026 5:13 UTC
2 points
0 comments13 min readLW link

Prac­ti­cal con­nec­tion with past lives

KatjaGrace2 Jul 2026 4:01 UTC
29 points
3 comments1 min readLW link
(worldspiritsockpuppet.substack.com)

Ca­reer Choice: Be­com­ing a Re­searcher in a Non-EA-Pri­or­ity Field vs Found­ing Tech Startup?

Master Chief2 Jul 2026 3:24 UTC
8 points
0 comments1 min readLW link