RSS

Jasmine Brazilek

Karma: 83

What is the Align­ment Com­mu­nity Think­ing?

6 Sep 2026 16:34 UTC
5 points
0 comments7 min readLW link

Re­search sci­en­tists for CaML: In­ves­ti­gat­ing whether mid-train­ing can sur­vive RL

Jasmine Brazilek1 Sep 2026 18:13 UTC
16 points
2 comments2 min readLW link

Co­er­cion and De­cep­tion in AI-to-AI Management

10 Aug 2026 16:13 UTC
12 points
0 comments8 min readLW link
(compassionalignedml.substack.com)

Com­mu­nity Polls on Align­ment Con­tro­ver­sies II

30 Jul 2026 21:05 UTC
9 points
15 comments2 min readLW link
(forum.effectivealtruism.org)

Would your AI travel agent book a bul­lfight? Test­ing whether agents con­sider an­i­mal welfare with­out be­ing prompted

17 Jul 2026 17:28 UTC
13 points
14 comments3 min readLW link

Assert, don’t de­scribe: how writ­ing style in train­ing data shapes an AI’s moral stance

17 Jun 2026 22:48 UTC
2 points
0 comments13 min readLW link
(forum.effectivealtruism.org)

[Linkpost] Com­mu­nity polls on al­ign­ment controversies

17 Jun 2026 0:09 UTC
8 points
7 comments1 min readLW link
(forum.effectivealtruism.org)

Doc­u­ment-tun­ing in­stills durable an­i­mal com­pas­sion in LLMs (and gen­er­al­izes to hu­mans)

21 May 2026 3:29 UTC
11 points
0 comments6 min readLW link