RSS

Venkat T

Karma: 35

Independent Technical AI safety researcher.

Alignment, Control, Interp, Evals.

A non-gen­er­a­tive model as a trusted mon­i­tor for AI Con­trol: Test­ing TypeSafe’s Jev

Venkat T18 Sep 2026 16:06 UTC
20 points
0 comments15 min readLW link

Venkat T’s Shortform

Venkat T17 Sep 2026 0:23 UTC
2 points
1 comment1 min readLW link

Train­ing against the mon­i­tor: What hap­pens dur­ing Obfus­cated Ad­ver­sar­ial Train­ing?

Venkat T9 Sep 2026 1:36 UTC
17 points
0 comments31 min readLW link