RSS

emanuelr

Karma: 117

Emanuel Ruzak

How much of ML re­search is about AI safety, what is it about, and who’s do­ing it?

15 Jul 2026 2:57 UTC
23 points
0 comments5 min readLW link

Nat­u­rally learned be­hav­iors in deep MLPs re­sist de­tec­tion by both hu­man and learned algorithms

emanuelr22 Jun 2026 9:10 UTC
17 points
0 comments12 min readLW link

An AI al­ign­ment re­search agenda based on asym­met­ric de­bate and mon­i­tor­ing.

emanuelr10 Apr 2026 6:23 UTC
4 points
0 comments19 min readLW link

Ex­plor­ing be­lief states in LLM chains of thought

emanuelr27 Sep 2025 1:09 UTC
6 points
2 comments7 min readLW link