RSS

Cam

Karma: 733

I help run www.geodesicresearch.org

Why study proto-train­ing gam­ing as an ad­ver­sar­ial al­ign­ment failure mode?

8 Jul 2026 17:07 UTC
55 points
0 comments7 min readLW link

Why study al­ign­ment in­ter­ven­tions on pre-RL check­points?

8 Jul 2026 17:07 UTC
65 points
2 comments6 min readLW link

An­nounc­ing Geodesic Research

27 May 2026 16:40 UTC
84 points
2 comments5 min readLW link

Learned Chain-of-Thought Obfus­ca­tion Gen­er­al­ises to Unseen Tasks

21 May 2026 10:11 UTC
31 points
0 comments5 min readLW link
(arxiv.org)

Align­ment Pre­train­ing: AI Dis­course Causes Self-Fulfilling (Mis)alignment

21 Dec 2025 0:53 UTC
207 points
25 comments9 min readLW link

Ar­chi­tec­tures for In­creased Ex­ter­nal­i­sa­tion of Reasoning

26 Nov 2025 20:24 UTC
38 points
2 comments13 min readLW link

Gen­er­al­i­sa­tion Hack­ing: a first look at ad­ver­sar­ial gen­er­al­i­sa­tion failures in de­liber­a­tive alignment

17 Nov 2025 21:44 UTC
54 points
2 comments8 min readLW link

Open-weight train­ing prac­tices and im­pli­ca­tions for CoT monitorability

4 Nov 2025 10:49 UTC
21 points
0 comments9 min readLW link

Cam’s Shortform

Cam9 Feb 2025 17:32 UTC
1 point
20 comments1 min readLW link