RSS

Jason R Brown

Karma: 617

A Con­cep­tual Frame­work for Rea­son­ing about Ex­plo­ra­tion Hacking

8 Sep 2026 19:13 UTC
38 points
0 comments17 min readLW link

Ex­plo­ra­tion Hack­ing in AI De­bate: Ini­tial Em­pirics and Gen­er­al­i­sa­tion Splitting

8 Sep 2026 19:13 UTC
33 points
0 comments13 min readLW link

A Post-Mortem for My Goal Crys­talli­sa­tion Project

17 Jul 2026 14:30 UTC
38 points
0 comments10 min readLW link

Op­ti­miser Choice Can Am­plify or Sup­press Emer­gent Misalignment

9 Jul 2026 10:00 UTC
63 points
2 comments4 min readLW link

Ja­son R Brown’s Shortform

Jason R Brown18 Jun 2026 14:26 UTC
4 points
2 comments1 min readLW link

Es­ti­mat­ing No-CoT Task-Com­ple­tion Time Hori­zons of Fron­tier AI Models

10 Jun 2026 17:58 UTC
280 points
23 comments4 min readLW link

Devel­op­men­tal Cog­ni­tive In­ter­pretabil­ity: A Re­search Agenda for Model­ling Gen­er­al­i­sa­tion and Pre­dict­ing Agent Behaviour

29 May 2026 9:56 UTC
70 points
0 comments7 min readLW link

AI Safety Re­search Futarchy: Us­ing Pre­dic­tion Mar­kets to Choose Re­search Pro­jects for MARS

Jason R Brown30 Sep 2025 15:37 UTC
37 points
10 comments4 min readLW link

TAMing The Align­ment Problem

Jason R Brown7 Apr 2025 8:47 UTC
13 points
2 comments11 min readLW link

Quan­tify­ing Gen­eral Intelligence

Jason R Brown17 Jun 2022 21:57 UTC
9 points
6 comments13 min readLW link