RSS

Alejandro Aristizabal

Karma: 85

Au­to­mated al­ign­ment runs are hard to study!

13 Aug 2026 15:04 UTC
53 points
3 comments9 min readLW link

De­bate with Self-Play Best-of-N Optimization

9 Jul 2026 15:29 UTC
49 points
2 comments14 min readLW link

Ex­plor­ing Shard-like Be­hav­ior: Em­piri­cal In­sights into Con­tex­tual De­ci­sion-Mak­ing in RL Agents

Alejandro Aristizabal29 Sep 2024 0:32 UTC
6 points
0 comments15 min readLW link