RSS

Jasmine Brazilek

Karma: 83

What is the Align­ment Com­mu­nity Think­ing?

6 Sep 2026 16:34 UTC
5 points
0 comments7 min readLW link

Re­search sci­en­tists for CaML: In­ves­ti­gat­ing whether mid-train­ing can sur­vive RL

Jasmine Brazilek1 Sep 2026 18:13 UTC
16 points
2 comments2 min readLW link

Co­er­cion and De­cep­tion in AI-to-AI Management

10 Aug 2026 16:13 UTC
12 points
0 comments8 min readLW link
(compassionalignedml.substack.com)

Com­mu­nity Polls on Align­ment Con­tro­ver­sies II

30 Jul 2026 21:05 UTC
9 points
15 comments2 min readLW link
(forum.effectivealtruism.org)

Would your AI travel agent book a bul­lfight? Test­ing whether agents con­sider an­i­mal welfare with­out be­ing prompted

17 Jul 2026 17:28 UTC
13 points
14 comments3 min readLW link

Assert, don’t de­scribe: how writ­ing style in train­ing data shapes an AI’s moral stance

17 Jun 2026 22:48 UTC
2 points
0 comments13 min readLW link
(forum.effectivealtruism.org)

[Linkpost] Com­mu­nity polls on al­ign­ment controversies

17 Jun 2026 0:09 UTC
8 points
7 comments1 min readLW link
(forum.effectivealtruism.org)