RSS

In­tent Is All You Need.

Not Sure24 Jul 2026 19:52 UTC
3 points
1 comment2 min readLW link

Stable Sys­tems Have Stable Outputs

Deixis24 Jul 2026 19:21 UTC
5 points
0 comments4 min readLW link

The AI In­dus­trial Ex­plo­sion — Part 5: Given AGI, au­tomat­ing phys­i­cal pro­duc­tion is prob­a­bly not that hard

djbinder24 Jul 2026 19:15 UTC
11 points
0 comments25 min readLW link
(defensesindepth.bio)

Where does hint-fol­low­ing and con­ceal­ment arise? A case study on OLMo-3 checkpoints

24 Jul 2026 19:14 UTC
9 points
0 comments4 min readLW link

Should we be wor­ried about how good AI is get­ting at cod­ing au­tonomous drones?

Lukas Petersson24 Jul 2026 16:44 UTC
6 points
1 comment1 min readLW link

LLMs are (still) mostly pow­ered by imi­ta­tive learn­ing, not RL

Steven Byrnes24 Jul 2026 14:26 UTC
51 points
12 comments9 min readLW link

Democ­racy isn’t ready for the AI revolution

Sophia Gore24 Jul 2026 14:17 UTC
25 points
3 comments5 min readLW link

Does dis­till­ing Claude carry the per­sona with it?

24 Jul 2026 12:31 UTC
27 points
2 comments10 min readLW link

Ge­or­gia Tech AI Safety Ini­ti­a­tive Ret­ro­spec­tive 2025-2026

24 Jul 2026 11:55 UTC
22 points
1 comment7 min readLW link

[Linkpost] Thoughts on the Re­cent OpenAI Hack

Linch24 Jul 2026 1:51 UTC
17 points
2 comments4 min readLW link

Should OpenAI’s rogue agent be pun­ished?

groblegark24 Jul 2026 1:20 UTC
1 point
3 comments1 min readLW link

Eval­u­at­ing Red Team and Blue Team Ca­pa­bil­ity for AI Con­trol Research

Ram Potham24 Jul 2026 1:11 UTC
9 points
0 comments8 min readLW link
(dearfutureais.substack.com)

Fix­ing re­wards for NLA to re­duce confabulation

SEONG PYO HONG24 Jul 2026 0:55 UTC
9 points
0 comments4 min readLW link

An­thropic’s J-Lens: A Re­search Eng­ineer’s Analysis

willkn24 Jul 2026 0:54 UTC
7 points
0 comments11 min readLW link

Con­tra Ge­orge Hotz on “AI 2040 and the Cult of In­tel­li­gence”

Matthew Tromp23 Jul 2026 22:43 UTC
8 points
0 comments8 min readLW link
(substack.com)

The Model Or­ganism Lot­tery: Model Or­ganism In­ter­pretabil­ity Strongly Depends on Train­ing Methodology

23 Jul 2026 22:37 UTC
11 points
0 comments6 min readLW link
(arxiv.org)

vibes-based think­ing as a cul­tural re­sponse to un­knowns

madelineberzak23 Jul 2026 22:35 UTC
7 points
0 comments8 min readLW link

Why an LLM can­not ac­cu­mu­late concepts

Zenya23 Jul 2026 21:40 UTC
−11 points
0 comments2 min readLW link

Red-team­ing LLM un­learn­ing: LUNAR’s “for­got­ten” knowl­edge is still recoverable

prefrontal23 Jul 2026 21:39 UTC
10 points
0 comments5 min readLW link

In­cep­tion in Diffu­sionGemma—Jailbreak­ing a Diffu­sion Lan­guage Model by Pin­ning To­kens Any­where on the Canvas

23 Jul 2026 21:39 UTC
15 points
0 comments11 min readLW link