RSS

VojtaKovarik

Karma: 1,064

My original background is in mathematics (analysis, topology, Banach spaces) and game theory (imperfect information games). Nowadays, I do AI alignment research (mostly systemic risks, sometimes pondering about “consequentionalist reasoning”).

Agenda: In­fras­truc­ture for Trad­ing with Par­tially Misal­igned AIs

VojtaKovarik7 Aug 2026 14:13 UTC
17 points
0 comments9 min readLW link
(limits-of-evaluation.org)

Ori­ent­ing Towards Over­sight: Which AIs Should Want to Defect?

31 Jul 2026 15:20 UTC
21 points
0 comments13 min readLW link
(limits-of-evaluation.org)

The Hu­man Sub­sti­tu­tion Test as a San­ity Check for AI Evaluations

10 Jul 2026 17:27 UTC
31 points
5 comments8 min readLW link
(limits-of-evaluation.org)

En­tan­gle­ment Between an AI and Its Environment

7 Jul 2026 13:28 UTC
27 points
2 comments11 min readLW link
(limits-of-evaluation.org)

De­ploy­ment Aware­ness Mat­ters More Than Eval­u­a­tion Awareness

26 Jun 2026 22:54 UTC
46 points
7 comments7 min readLW link
(limits-of-evaluation.org)