Questions
Events
Shortform
Alignment Forum
AF Comments
Home
Featured
All
Tags
Recent
Comments
Archive
Sequences
About
Search
Log In
All
2005
2006
2007
2008
2009
2010
2011
2012
2013
2014
2015
2016
2017
2018
2019
2020
2021
2022
2023
2024
2025
2026
All
Jan
Feb
Mar
Apr
May
Jun
Jul
Aug
Sep
Oct
Nov
Dec
Page
1
SolidGoldMagikarp (plus, prompt generation)
Jessica Rumbelow
and
mwatkins
5 Feb 2023 22:02 UTC
680
points
208
comments
12
min read
LW
link
1
review
The Waluigi Effect (mega-post)
Cleo Nardo
3 Mar 2023 3:22 UTC
654
points
188
comments
16
min read
LW
link
The Talk: a brief explanation of sexual dimorphism
Malmesbury
18 Sep 2023 16:23 UTC
555
points
79
comments
16
min read
LW
link
3
reviews
How much do you believe your results?
Eric Neyman
6 May 2023 20:31 UTC
531
points
19
comments
15
min read
LW
link
4
reviews
(ericneyman.wordpress.com)
Focus on the places where you feel shocked everyone’s dropping the ball
So8res
2 Feb 2023 0:27 UTC
518
points
65
comments
4
min read
LW
link
3
reviews
The ants and the grasshopper
Richard_Ngo
4 Jun 2023 22:00 UTC
517
points
45
comments
5
min read
LW
link
4
reviews
(www.narrativeark.xyz)
Significantly Enhancing Adult Intelligence With Gene Editing May Be Possible
GeneSmith
and
kman
12 Dec 2023 18:14 UTC
475
points
212
comments
33
min read
LW
link
2
reviews
Things I Learned by Spending Five Thousand Hours In Non-EA Charities
jenn
1 Jun 2023 20:48 UTC
451
points
37
comments
8
min read
LW
link
1
review
(jenn.site)
Please don’t throw your mind away
TsviBT
15 Feb 2023 21:41 UTC
444
points
50
comments
18
min read
LW
link
1
review
Steering GPT-2-XL by adding an activation vector
TurnTrout
,
Monte M
,
David Udell
,
lisathiergart
and
Ulisse Mini
13 May 2023 18:42 UTC
442
points
98
comments
50
min read
LW
link
1
review
GPTs are Predictors, not Imitators
Eliezer Yudkowsky
8 Apr 2023 19:59 UTC
435
points
100
comments
3
min read
LW
link
3
reviews
Douglas Hofstadter changes his mind on Deep Learning & AI risk (June 2023)?
gwern
3 Jul 2023 0:48 UTC
432
points
54
comments
7
min read
LW
link
(www.youtube.com)
Bing Chat is blatantly, aggressively misaligned
evhub
15 Feb 2023 5:29 UTC
397
points
181
comments
2
min read
LW
link
1
review
Social Dark Matter
Duncan Sabien (Inactive)
16 Nov 2023 20:00 UTC
390
points
133
comments
34
min read
LW
link
2
reviews
Statement on AI Extinction—Signed by AGI Labs, Top Academics, and Many Other Notable Figures
Dan H
30 May 2023 9:05 UTC
383
points
78
comments
1
min read
LW
link
1
review
(www.safe.ai)
Noting an error in Inadequate Equilibria
Matthew Barnett
8 Feb 2023 1:33 UTC
379
points
60
comments
2
min read
LW
link
2
reviews
How it feels to have your mind hacked by an AI
blaked
12 Jan 2023 0:33 UTC
377
points
222
comments
17
min read
LW
link
How to have Polygenically Screened Children
GeneSmith
7 May 2023 16:01 UTC
374
points
128
comments
27
min read
LW
link
1
review
My Objections to “We’re All Gonna Die with Eliezer Yudkowsky”
Quintin Pope
21 Mar 2023 0:06 UTC
372
points
233
comments
39
min read
LW
link
1
review
Fucking Goddamn Basics of Rationalist Discourse
LoganStrohl
4 Feb 2023 1:47 UTC
365
points
104
comments
1
min read
LW
link
3
reviews
Childhoods of exceptional people
Henrik Karlsson
6 Feb 2023 17:27 UTC
354
points
62
comments
15
min read
LW
link
1
review
(escapingflatland.substack.com)
Guide to rationalist interior decorating
mingyuan
19 Jun 2023 6:47 UTC
352
points
53
comments
12
min read
LW
link
4
reviews
Shallow review of live agendas in alignment & safety
technicalities
and
Stag
27 Nov 2023 11:10 UTC
351
points
73
comments
29
min read
LW
link
1
review
Cyborgism
Niki Dupuis
and
janus
10 Feb 2023 14:47 UTC
342
points
47
comments
35
min read
LW
link
2
reviews
Inside Views, Impostor Syndrome, and the Great LARP
johnswentworth
25 Sep 2023 16:08 UTC
341
points
54
comments
5
min read
LW
link
Shutting Down the Lightcone Offices
habryka
and
Ben Pace
14 Mar 2023 22:47 UTC
340
points
104
comments
17
min read
LW
link
2
reviews
Model Organisms of Misalignment: The Case for a New Pillar of Alignment Research
evhub
,
Nicholas Schiefer
,
Carson Denison
and
Ethan Perez
8 Aug 2023 1:30 UTC
338
points
30
comments
18
min read
LW
link
1
review
Against Almost Every Theory of Impact of Interpretability
Charbel-Raphaël
17 Aug 2023 18:44 UTC
337
points
93
comments
26
min read
LW
link
2
reviews
Understanding and controlling a maze-solving policy network
TurnTrout
,
peligrietzer
,
Ulisse Mini
,
Monte M
and
David Udell
11 Mar 2023 18:59 UTC
336
points
28
comments
23
min read
LW
link
When do “brains beat brawn” in Chess? An experiment
titotal
28 Jun 2023 13:33 UTC
335
points
107
comments
7
min read
LW
link
2
reviews
(titotal.substack.com)
EA Vegan Advocacy is not truthseeking, and it’s everyone’s problem
Elizabeth
28 Sep 2023 23:30 UTC
334
points
250
comments
22
min read
LW
link
2
reviews
(acesounderglass.com)
Book Review: How Minds Change
bc4026bd4aaa5b7fe
25 May 2023 17:55 UTC
329
points
53
comments
15
min read
LW
link
Sharing Information About Nonlinear
Ben Pace
7 Sep 2023 6:51 UTC
324
points
324
comments
34
min read
LW
link
The Parable of the King and the Random Process
moridinamael
1 Mar 2023 22:18 UTC
318
points
26
comments
6
min read
LW
link
3
reviews
Speaking to Congressional staffers about AI risk
Orpheus16
and
hath
4 Dec 2023 23:08 UTC
315
points
25
comments
15
min read
LW
link
1
review
Alignment Grantmaking is Funding-Limited Right Now
johnswentworth
19 Jul 2023 16:49 UTC
312
points
68
comments
1
min read
LW
link
On not getting contaminated by the wrong obesity ideas
Natália
28 Jan 2023 20:18 UTC
310
points
69
comments
30
min read
LW
link
LW Team is adjusting moderation policy
Raemon
4 Apr 2023 20:41 UTC
307
points
185
comments
3
min read
LW
link
AI Timelines
habryka
,
Daniel Kokotajlo
,
Ajeya Cotra
and
Ege Erdil
10 Nov 2023 5:28 UTC
303
points
147
comments
51
min read
LW
link
2
reviews
Deep Deceptiveness
So8res
21 Mar 2023 2:51 UTC
299
points
61
comments
14
min read
LW
link
1
review
Predictable updating about AI risk
Joe Carlsmith
8 May 2023 21:53 UTC
298
points
25
comments
36
min read
LW
link
1
review
The 101 Space You Will Always Have With You
Screwtape
29 Nov 2023 4:56 UTC
297
points
23
comments
6
min read
LW
link
1
review
Notes on Teaching in Prison
jsd
19 Apr 2023 1:53 UTC
295
points
13
comments
12
min read
LW
link
Pausing AI Developments Isn’t Enough. We Need to Shut it All Down by Eliezer Yudkowsky
jacquesthibs
29 Mar 2023 23:16 UTC
294
points
297
comments
3
min read
LW
link
(time.com)
Accidentally Load Bearing
jefftk
13 Jul 2023 16:10 UTC
292
points
19
comments
1
min read
LW
link
1
review
(www.jefftk.com)
Basics of Rationalist Discourse
Duncan Sabien (Inactive)
27 Jan 2023 2:40 UTC
291
points
193
comments
31
min read
LW
link
4
reviews
The 6D effect: When companies take risks, one email can be very powerful.
scasper
4 Nov 2023 20:08 UTC
290
points
43
comments
3
min read
LW
link
Hooray for stepping out of the limelight
So8res
1 Apr 2023 2:45 UTC
289
points
26
comments
1
min read
LW
link
Towards Monosemanticity: Decomposing Language Models With Dictionary Learning
Zac Hatfield-Dodds
5 Oct 2023 21:01 UTC
289
points
22
comments
2
min read
LW
link
1
review
(transformer-circuits.pub)
We don’t trade with ants
KatjaGrace
10 Jan 2023 23:50 UTC
283
points
110
comments
7
min read
LW
link
1
review
(worldspiritsockpuppet.com)
Back to top
Next