Archive
Sequences
About
Search
Log In
Questions
Events
Shortform
Alignment Forum
AF Comments
Home
Featured
All
Tags
Recent
Comments
RSS
Ram Potham
Karma:
220
Prev High Stakes Control at Redwood Research
(xkcd meme)
All
Posts
Comments
New
Top
Old
Evaluating Red Team and Blue Team Capability for AI Control Research
Ram Potham
24 Jul 2026 1:11 UTC
11
points
0
comments
8
min read
LW
link
(dearfutureais.substack.com)
Studying the role of Sandboxing for AI Control
Ram Potham
17 Jul 2026 19:05 UTC
12
points
0
comments
10
min read
LW
link
Making Credible Deals With AI
Ram Potham
14 Jul 2026 1:14 UTC
13
points
17
comments
17
min read
LW
link
(dearfutureais.substack.com)
Should we combine protocols for AI Control Research?
Ram Potham
and
Tyler Tracy
26 Jun 2026 19:02 UTC
9
points
0
comments
9
min read
LW
link
Scaffolding vs Reinforcement Finetuning for AI Forecasting
Ram Potham
30 Apr 2026 2:51 UTC
15
points
0
comments
4
min read
LW
link
Smart Contracts as Credible Commitments for Trading with Scheming AIs
Ram Potham
22 Apr 2026 15:46 UTC
6
points
0
comments
9
min read
LW
link
Introducing LinuxArena
Tyler Tracy
,
Ram Potham
,
Nick Kuhn
and
Myles H
20 Apr 2026 22:00 UTC
84
points
2
comments
4
min read
LW
link
Attack Selection In Agentic AI Control Evals Can Decrease Safety
Cath Ge-Wang
,
Tyler Crosse
,
hadad
,
Ram Potham
and
Tyler Tracy
14 Apr 2026 18:02 UTC
26
points
4
comments
18
min read
LW
link
I Tested LLM Agents on Simple Safety Rules. They Failed in Surprising and Informative Ways.
Ram Potham
25 Jun 2025 21:39 UTC
9
points
12
comments
6
min read
LW
link
AI Control Methods Literature Review
Ram Potham
18 Apr 2025 21:15 UTC
12
points
1
comment
9
min read
LW
link
Ram Potham’s Shortform
Ram Potham
23 Mar 2025 15:08 UTC
1
point
15
comments
1
min read
LW
link
Back to top