RSS

jsteinhardt

Karma: 6,307

User aware­ness in fron­tier models

6 Aug 2026 20:43 UTC
58 points
1 comment12 min readLW link
(transluce.org)

Foun­da­tion Models for Oversight

jsteinhardt28 Jul 2026 16:30 UTC
62 points
3 comments25 min readLW link
(bounded-regret.ghost.io)

Cog­ni­tive Se­cu­rity as an AI Safety Cause Area

jsteinhardt25 May 2026 18:30 UTC
158 points
20 comments2 min readLW link

The Case for Eval­u­at­ing Model Behaviors

jsteinhardt20 May 2026 18:42 UTC
42 points
3 comments3 min readLW link

Build­ing Tech­nol­ogy to Drive AI Governance

jsteinhardt18 Feb 2026 18:30 UTC
63 points
4 comments10 min readLW link
(bounded-regret.ghost.io)

Over­sight As­sis­tants: Turn­ing Com­pute into Understanding

jsteinhardt6 Jan 2026 0:50 UTC
85 points
7 comments9 min readLW link
(bounded-regret.ghost.io)