RSS

draganover

Karma: 478

Can we find whether mod­els have been back­doored?

6 Jul 2026 23:45 UTC
22 points
1 comment8 min readLW link

Your Model Or­ganisms Might Be Fried

18 Jun 2026 16:18 UTC
104 points
9 comments7 min readLW link

Learn­ings from start­ing an AI safety re­search team

5 Jun 2026 16:27 UTC
103 points
7 comments6 min readLW link

A Re­search Agenda for Se­cret Loyalties

13 May 2026 17:34 UTC
39 points
5 comments3 min readLW link

Why did peo­ple miss the point on Mythos?

draganover26 Apr 2026 12:15 UTC
48 points
14 comments5 min readLW link