RSS

Maheep Chaudhary

Karma: 57

Will AI Agents Pay to Avoid Killing An­i­mals?

30 Sep 2026 1:30 UTC
4 points
4 comments11 min readLW link

What is the Align­ment Com­mu­nity Think­ing?

6 Sep 2026 16:34 UTC
5 points
0 comments7 min readLW link

Co­er­cion and De­cep­tion in AI-to-AI Management

10 Aug 2026 16:13 UTC
12 points
0 comments8 min readLW link
(compassionalignedml.substack.com)

Would your AI travel agent book a bul­lfight? Test­ing whether agents con­sider an­i­mal welfare with­out be­ing prompted

17 Jul 2026 17:28 UTC
13 points
14 comments3 min readLW link

Aware­ness Jailbreak­ing: Re­veal­ing True Align­ment in Eval­u­a­tion-Aware Models

Maheep Chaudhary29 Dec 2025 21:29 UTC
11 points
0 comments4 min readLW link

Eval­u­a­tion Aware­ness Scales Pre­dictably in Open-Weights Large Lan­guage Models

Maheep Chaudhary19 Dec 2025 2:47 UTC
23 points
0 comments6 min readLW link