Archive
Sequences
About
Search
Log In
Questions
Events
Shortform
Alignment Forum
AF Comments
Home
Featured
All
Tags
Recent
Comments
RSS
New
Hot
Active
Old
Page
1
Misaligned AIs could use killer robots to take over
Omar Khursheed
and
TurnTrout
11 Aug 2026 19:03 UTC
7
points
0
comments
5
min read
LW
link
Measuring Spurious Correlations with Feature Strength
egan
11 Aug 2026 18:05 UTC
13
points
0
comments
10
min read
LW
link
AI governance work needs much better monitoring
jackultraphil
11 Aug 2026 17:42 UTC
3
points
0
comments
8
min read
LW
link
(fundinganthropalypse.com)
LLMs Are Starting To Noticeably Accelerate Our Work
johnswentworth
11 Aug 2026 17:06 UTC
63
points
3
comments
2
min read
LW
link
Software Is Not Soft
cylonator
11 Aug 2026 16:26 UTC
5
points
0
comments
1
min read
LW
link
How risky would it be to make powerful AI obey one or a few people?
cousin_it
and
Seth Herd
11 Aug 2026 16:15 UTC
45
points
0
comments
8
min read
LW
link
Extreme concentration of power over ASI has non-obvious advantages
Seth Herd
11 Aug 2026 16:13 UTC
19
points
0
comments
19
min read
LW
link
Those Who Make History
Raelifin
11 Aug 2026 13:59 UTC
21
points
1
comment
8
min read
LW
link
(open.substack.com)
Seeing things through in the age of AI
alkjash
11 Aug 2026 13:14 UTC
22
points
3
comments
1
min read
LW
link
Productive Signaling: Competitive Software Development, Not Competitive Programming
Adam Chlipala
11 Aug 2026 12:16 UTC
9
points
0
comments
9
min read
LW
link
The Next Ecology
Eigenbraid
11 Aug 2026 8:29 UTC
7
points
2
comments
5
min read
LW
link
Redux: (∃ Stochastic Natural Latent) Implies (∃ Deterministic Natural Latent)
David Lorell
11 Aug 2026 5:52 UTC
81
points
11
comments
2
min read
LW
link
On using crises to shift political will for AI
clickyquack
11 Aug 2026 4:54 UTC
9
points
0
comments
4
min read
LW
link
Revived Lightweight Transit Predictions Page
jefftk
11 Aug 2026 2:31 UTC
10
points
0
comments
1
min read
LW
link
(www.jefftk.com)
Before We Defer Research to AI: Measuring Apparent-Success-Seeking
Keira Leal
11 Aug 2026 2:21 UTC
12
points
3
comments
5
min read
LW
link
A Topic Detector, Not a Lie Detector: what J-space monitoring actually tracks
Melchior de Polignac
11 Aug 2026 2:18 UTC
10
points
0
comments
4
min read
LW
link
Models inherit the writer, not who the writer was imitating
0Chris5R
11 Aug 2026 2:16 UTC
7
points
0
comments
10
min read
LW
link
Creative math research by AI as the latest sign of the end
Mitchell_Porter
11 Aug 2026 1:37 UTC
45
points
9
comments
2
min read
LW
link
A study on instability of LLM responses as a behavioral signature of self-Referential reports.
PARAS BALANI
11 Aug 2026 1:01 UTC
7
points
0
comments
1
min read
LW
link
Q: Is dual-use alignment-complete problem?
kapedalex
10 Aug 2026 21:40 UTC
10
points
0
comments
1
min read
LW
link
Back to top
Next