Lazy Has­selback Pommes Anna

Brendan Long26 Jan 2025 21:30 UTC
16 points
19 comments3 min readLW link
(www.brendanlong.com)

Links and short notes, 2025-01-26: At­las Shrugged and the ir­re­place­able founder, pump­ing sta­tions and civic pride, and thoughts on the eve of AGI

jasoncrawford26 Jan 2025 20:52 UTC
8 points
1 comment3 min readLW link
(newsletter.rootsofprogress.org)

The gen­er­al­iza­tion phase diagram

Dmitry Vaintrob26 Jan 2025 20:30 UTC
28 points
2 comments16 min readLW link

Start­ing an Egan High School

Chris Wintergreen26 Jan 2025 19:02 UTC
9 points
2 comments6 min readLW link

Nar­ra­tives as cat­a­lysts of catas­trophic trajectories

EQ26 Jan 2025 19:01 UTC
6 points
0 comments8 min readLW link
(eqmind.substack.com)

The many failure modes of con­sumer-grade LLMs

dereshev26 Jan 2025 19:01 UTC
2 points
0 comments8 min readLW link

so you have a chronic health issue

agencypilled26 Jan 2025 19:00 UTC
23 points
10 comments5 min readLW link

If you wanted to ac­tu­ally re­duce the trade deficit, how would you do it?

Logan Zoellner26 Jan 2025 18:04 UTC
6 points
5 comments2 min readLW link

Anatomy of a Dance Class: A step by step guide

Nathan Young26 Jan 2025 18:02 UTC
13 points
0 comments10 min readLW link
(nathanpmyoung.substack.com)

Why care about AI per­son­hood?

Francis Rhys Ward26 Jan 2025 11:24 UTC
43 points
6 comments3 min readLW link

Nav­i­gat­ing Diver­sity: Un­der­stand­ing Hu­man Be­hav­iors Through Ge­net­ics, Neu­ro­di­ver­gence, and Trauma

j_passeri26 Jan 2025 8:23 UTC
1 point
0 comments3 min readLW link

Kessler’s Se­cond Syndrome

Jesse Hoogland26 Jan 2025 7:04 UTC
70 points
2 comments3 min readLW link

[Question] En­hanced Clar­ity to Bridge the AI La­bel­ing Gap?

Pathways26 Jan 2025 6:48 UTC
1 point
0 comments1 min readLW link

Brainrot

Jesse Hoogland26 Jan 2025 5:35 UTC
43 points
0 comments3 min readLW link

Notes on Argentina

Annapurna26 Jan 2025 3:51 UTC
18 points
5 comments4 min readLW link
(jorgevelez.substack.com)

[Question] Recom­men­da­tions for Re­cent Posts/​Se­quences on In­stru­men­tal Ra­tion­al­ity?

Benjamin Hendricks26 Jan 2025 0:41 UTC
13 points
3 comments1 min readLW link

Ano­ma­lous To­kens in Deep­Seek-V3 and r1

henry25 Jan 2025 22:55 UTC
145 points
3 comments7 min readLW link

The Ris­ing Sea

Jesse Hoogland25 Jan 2025 20:48 UTC
97 points
9 comments2 min readLW link

Monet: Mix­ture of Monose­man­tic Ex­perts for Trans­form­ers Explained

CalebMaresca25 Jan 2025 19:37 UTC
31 points
2 comments11 min readLW link

AI and Non-Ex­is­tence.

Eleven25 Jan 2025 19:36 UTC
−3 points
9 comments2 min readLW link

Agents don’t have to be al­igned to help us achieve an in­definite pause.

Hastings25 Jan 2025 18:51 UTC
31 points
0 comments3 min readLW link

[Question] AI Safety in secret

Michael Flood25 Jan 2025 18:16 UTC
7 points
0 comments1 min readLW link

On polytopes

Dmitry Vaintrob25 Jan 2025 13:56 UTC
56 points
5 comments12 min readLW link

At­tri­bu­tion-based pa­ram­e­ter decomposition

25 Jan 2025 13:12 UTC
109 points
21 comments4 min readLW link
(publications.apolloresearch.ai)

A con­cise defi­ni­tion of what it means to win

testingthewaters25 Jan 2025 6:37 UTC
4 points
1 comment5 min readLW link
(aclevername.substack.com)

[Question] A Float­ing Cube—Re­jected HLE submission

Shankar Sivarajan25 Jan 2025 4:52 UTC
8 points
1 comment1 min readLW link

Why I’m Pour­ing Cold Water in My Left Ear, and You Should Too

Maloew24 Jan 2025 23:13 UTC
12 points
0 comments2 min readLW link

Coun­ter­in­tu­itive effects of min­i­mum prices

dynomight24 Jan 2025 23:05 UTC
25 points
0 comments8 min readLW link
(dynomight.net)

AXRP Epi­sode 38.6 - Joel Lehman on Pos­i­tive Vi­sions of AI

DanielFilan24 Jan 2025 23:00 UTC
10 points
0 comments9 min readLW link

Lo­cat­ing and Edit­ing Knowl­edge in LMs

Dhananjay Ashok24 Jan 2025 22:53 UTC
1 point
0 comments4 min readLW link

How are Those AI Par­ti­ci­pants Do­ing Any­way?

mushroomsoup24 Jan 2025 22:37 UTC
4 points
0 comments10 min readLW link

Six Thoughts on AI Safety

Boaz Barak24 Jan 2025 22:20 UTC
95 points
56 comments15 min readLW link

In­stru­men­tal Goals Are A Differ­ent And Friendlier Kind Of Thing Than Ter­mi­nal Goals

24 Jan 2025 20:20 UTC
210 points
62 comments5 min readLW link

Yud­kowsky on The Tra­jec­tory podcast

Seth Herd24 Jan 2025 19:52 UTC
71 points
39 comments2 min readLW link
(www.youtube.com)

Em­piri­cal In­sights into Fea­ture Geom­e­try in Sparse Autoencoders

Jason Boxi Zhang24 Jan 2025 19:02 UTC
7 points
0 comments11 min readLW link

Liron Shapira vs Ken Stan­ley on Doom De­bates. A review

TheManxLoiner24 Jan 2025 18:01 UTC
10 points
0 comments14 min readLW link

Is there such a thing as an im­pos­si­ble pro­tein?

Abhishaike Mahajan24 Jan 2025 17:12 UTC
15 points
3 comments4 min readLW link
(www.owlposting.com)

Star­gate AI-1

Zvi24 Jan 2025 15:20 UTC
85 points
1 comment18 min readLW link
(thezvi.wordpress.com)

QFT and neu­ral nets: the ba­sic idea

Dmitry Vaintrob24 Jan 2025 13:54 UTC
28 points
0 comments8 min readLW link

Elic­it­ing bad contexts

24 Jan 2025 10:39 UTC
37 points
9 comments3 min readLW link

In­sights from “The Manga Guide to Phys­iol­ogy”

TurnTrout24 Jan 2025 5:18 UTC
27 points
3 comments1 min readLW link
(turntrout.com)

[Question] Do you con­sider perfect surveillance in­evitable?

samuelshadrach24 Jan 2025 4:57 UTC
17 points
34 comments1 min readLW link

Un­con­trol­lable: A Sur­pris­ingly Good In­tro­duc­tion to AI Risk

PeterMcCluskey24 Jan 2025 4:30 UTC
16 points
1 comment1 min readLW link
(bayesianinvestor.com)

Con­tra Dances Get­ting Shorter and Earlier

jefftk23 Jan 2025 23:30 UTC
11 points
0 comments2 min readLW link
(www.jefftk.com)

Start­ing Thoughts on RLHF

Michael Flood23 Jan 2025 22:16 UTC
2 points
0 comments5 min readLW link

Up­dat­ing and Edit­ing Fac­tual Knowl­edge in Lan­guage Models

Dhananjay Ashok23 Jan 2025 19:34 UTC
2 points
2 comments10 min readLW link

AI com­pa­nies are un­likely to make high-as­surance safety cases if timelines are short

ryan_greenblatt23 Jan 2025 18:41 UTC
146 points
5 comments13 min readLW link

AISN #46: The Transition

23 Jan 2025 18:09 UTC
8 points
0 comments5 min readLW link
(newsletter.safe.ai)

What does suc­cess look like?

Raymond Douglas23 Jan 2025 17:48 UTC
11 points
0 comments3 min readLW link

AI #100: Meet the New Boss

Zvi23 Jan 2025 15:40 UTC
50 points
4 comments69 min readLW link
(thezvi.wordpress.com)