Fun­da­men­tal Uncer­tainty Es­say Contestants

Gordon Seidoh Worley5 Aug 2026 23:01 UTC
15 points
0 comments1 min readLW link
(www.uncertainupdates.com)

Alex Turner on Leav­ing Google Deep­Mind and Disagree­ments with Yudkowsky

Liron5 Aug 2026 22:45 UTC
59 points
1 comment40 min readLW link

Thomas Schel­ling’s No­bel Prize Speech: An As­ton­ish­ing Sixty Years: The Le­gacy of Hiroshima

Nathan Young5 Aug 2026 21:30 UTC
26 points
0 comments19 min readLW link

Ter­mi­nal-Bench Leader­board Rank­ings: Luck or Skill?

nateg5510155 Aug 2026 21:20 UTC
7 points
0 comments2 min readLW link

An ar­gu­ment of Parfit’s re­con­sid­ered with log­i­cal de­ci­sion theory

transhumanist_atom_understander5 Aug 2026 20:47 UTC
12 points
14 comments5 min readLW link

R-lens: Mak­ing J-lens More Faith­ful on Early Layers

5 Aug 2026 20:02 UTC
83 points
5 comments7 min readLW link

Mea­sur­ing cod­ing agent mis­al­ign­ment in the wild

snaz5 Aug 2026 20:00 UTC
40 points
2 comments7 min readLW link

An In­ter­na­tional AI Slow­down Is Ready When­ever Poli­ti­ci­ans Are

5 Aug 2026 19:15 UTC
56 points
4 comments9 min readLW link
(newsletter.ai-frontiers.org)

Ar­gu­ments for P

Cleo Nardo5 Aug 2026 19:00 UTC
549 points
66 comments2 min readLW link

Gen­er­al­ized athe­ism rules out “in­ac­cu­rate simu­la­tion”-ism.

Eliezer Yudkowsky5 Aug 2026 18:35 UTC
165 points
56 comments9 min readLW link

Berkeley Ge­nomics Pro­ject seek­ing hires (and col­labs)

TsviBT5 Aug 2026 18:32 UTC
41 points
9 comments9 min readLW link

Charles Good­hart Ele­men­tary School

Zack_M_Davis5 Aug 2026 17:06 UTC
45 points
3 comments2 min readLW link
(zackmdavis.net)

The Three AI Pills

Zvi5 Aug 2026 16:10 UTC
75 points
22 comments15 min readLW link
(thezvi.wordpress.com)

Notes on the pos­si­bil­ity of moral progress

MichaelDickens5 Aug 2026 16:05 UTC
14 points
10 comments7 min readLW link

Re­grant­ing in 2026: now more than ever

5 Aug 2026 13:59 UTC
2 points
0 comments5 min readLW link
(manifund.substack.com)

Ran­dom­ness and Dooms­day for Weak agents

Ben5 Aug 2026 12:24 UTC
32 points
6 comments7 min readLW link

Ver­ti­cal Tabs in Chrome

jefftk5 Aug 2026 12:10 UTC
48 points
3 comments1 min readLW link
(www.jefftk.com)

The goal­posts are shrouded, not moving

philh5 Aug 2026 11:40 UTC
134 points
6 comments1 min readLW link
(reasonableapproximation.net)

Help (re)start AI Safety at UPenn!

5 Aug 2026 9:42 UTC
15 points
0 comments1 min readLW link

Per­sona Cor­rup­tion and Role Mis­cast­ing in Emer­gent Misalignment

unruly abstractions5 Aug 2026 4:56 UTC
17 points
0 comments12 min readLW link

Warn­ing Shots for AI Ex­is­ten­tial Risk: Will so­ciety an­swer the wake-up calls it re­ceives?

5 Aug 2026 4:30 UTC
31 points
0 comments13 min readLW link
(techgov.intelligence.org)

AISafety.com Hackathon 2026

Bryce Robertson5 Aug 2026 4:15 UTC
4 points
0 comments1 min readLW link

See all up­com­ing AI safety events and train­ing programs

5 Aug 2026 3:28 UTC
19 points
0 comments1 min readLW link

Var­i­ous Ways to En­able People

warner5 Aug 2026 0:03 UTC
12 points
0 comments2 min readLW link

The AI Race is Not a Pri­soner’s Dilemma

Vaughn Papenhausen4 Aug 2026 23:45 UTC
67 points
6 comments8 min readLW link

Norms Are Bro­ken Laws

sirawit4 Aug 2026 23:42 UTC
−14 points
5 comments4 min readLW link

Why I think we live in a “simu­la­tion”

Eye You4 Aug 2026 22:43 UTC
4 points
6 comments5 min readLW link

Re­turn­ing to ARC

paulfchristiano4 Aug 2026 22:27 UTC
387 points
36 comments9 min readLW link

Geo­met­ric Ra­tion­al­ity acts lin­early in ad­di­tive scenarios

Bunthut4 Aug 2026 22:04 UTC
18 points
0 comments2 min readLW link

Most donors get risk wrong

jackultraphil4 Aug 2026 20:35 UTC
14 points
0 comments10 min readLW link
(open.substack.com)

An Ami­ca­ble Ape

James Stephen Brown4 Aug 2026 19:49 UTC
−1 points
0 comments4 min readLW link
(nonzerosum.games)

Beyond Tech’s Power: On Sta­tus, Sacri­fice, and the Search for Legitimacy

mhdempsey4 Aug 2026 19:46 UTC
3 points
0 comments10 min readLW link

Why don’t we just give AI the an­swers?

Brendan Long4 Aug 2026 19:46 UTC
89 points
29 comments2 min readLW link

An­nounc­ing Lat­eral Work­shop for ex­pe­rienced pro­fes­sion­als mov­ing into AI safety

4 Aug 2026 19:00 UTC
38 points
1 comment5 min readLW link
(forum.effectivealtruism.org)

Does Your LLM Trust You?

Bhalewow4 Aug 2026 18:57 UTC
8 points
0 comments11 min readLW link

Don’t Dither

sarahconstantin4 Aug 2026 17:10 UTC
41 points
10 comments4 min readLW link
(sarahconstantin.substack.com)

There Will Come Soft Rains

tanagrabeast4 Aug 2026 16:56 UTC
130 points
5 comments3 min readLW link

Com­mod­ify­ing Thinking

Rex Heng4 Aug 2026 16:55 UTC
2 points
1 comment7 min readLW link

Would We See It Com­ing? Prefer­ence Falsifi­ca­tion Cas­cades in Multi-Agent Systems

Sophia Hatz4 Aug 2026 16:35 UTC
15 points
0 comments10 min readLW link

Lets fix fac­tory farm­ing with AI

Kate Delbeke4 Aug 2026 15:28 UTC
8 points
0 comments3 min readLW link

Why For­ma­tion Re­search is Work­ing on Se­cret Loyalties

Alfie Lamerton4 Aug 2026 14:52 UTC
23 points
0 comments3 min readLW link

Gen­eral ca­pa­bil­ity—and ca­pa­bil­ities gen­er­ally—have no good y-axis

Thrasymachus4 Aug 2026 14:25 UTC
40 points
9 comments31 min readLW link
(forum.effectivealtruism.org)

Rewrite All the Code, All the Time

Adam Chlipala4 Aug 2026 13:20 UTC
31 points
15 comments8 min readLW link

When should we trust a la­tent rep­re­sen­ta­tion?

Ratnaditya J4 Aug 2026 11:16 UTC
8 points
1 comment8 min readLW link

Work­ing on Eco­nomics with Fable 5

Wilsoniumite4 Aug 2026 8:10 UTC
5 points
4 comments1 min readLW link

What if peo­ple value their work mat­ter­ing?

Tim H4 Aug 2026 6:02 UTC
31 points
9 comments4 min readLW link

[LINK] The Fer­rett fails will save against the Dark Arts

CronoDAS4 Aug 2026 5:29 UTC
22 points
1 comment1 min readLW link
(theferrett.substack.com)

At­tack­ers Can Sublimi­nally Im­plant a Back­door at Low Sam­ple Count Without Prompt Access

keshavs3 Aug 2026 22:08 UTC
23 points
3 comments5 min readLW link

Selec­tive Identity

warner3 Aug 2026 21:38 UTC
12 points
0 comments2 min readLW link

III. An­thropic rea­son­ing has is­sues with in­finite wor­lds; D-SIA can fix this

Stuart_Armstrong3 Aug 2026 21:31 UTC
18 points
5 comments17 min readLW link