Archive
Sequences
About
Search
Log In
Questions
Events
Shortform
Alignment Forum
AF Comments
Home
Featured
All
Tags
Recent
Comments
RSS
Multi-Agent Safety
Tag
Last edit:
9 Feb 2026 12:45 UTC
by
Hiroshi Yamakawa
Relevant
New
Old
The Inheritance Threshold: do behavioral norms survive a handoff between LLM agents?
Viktor Trncik
13 Jul 2026 17:32 UTC
1
point
0
comments
3
min read
LW
link
(doi.org)
The Multi-Agent Minefield: Can LLMs Cooperate to Avoid Global Catastrophe?
Isabel Dahlgren
16 Feb 2026 18:53 UTC
1
point
0
comments
5
min read
LW
link
Coercion and Deception in AI-to-AI Management
jonahmattwoodward
,
Jasmine Brazilek
,
MilesTS
and
Maheep Chaudhary
10 Aug 2026 16:13 UTC
12
points
0
comments
8
min read
LW
link
(compassionalignedml.substack.com)
I’m 18, Failed Chemistry, and I Think I Found Something in the Alignment Problem
Vansh Ahuja
16 May 2026 20:25 UTC
1
point
0
comments
1
min read
LW
link
A Multi-Agent Extension for Petri
carissacullen
22 Jul 2026 21:51 UTC
10
points
0
comments
4
min read
LW
link
Haning Alignment Protocol: Emergent Human-Compatible Values in Hybrid Multi-Agent Environments (A Conceptual Proposal)
Josh Haning
25 Jun 2026 20:45 UTC
1
point
0
comments
1
min read
LW
link
The Multi-Agent Minefield: Can LLMs Cooperate to Avoid Global Catastrophe?
Zhijing Jin
,
Thao Pham
,
TerryJCZhang
,
pepijn_cobben
,
Angelo Huang
,
Isabel Dahlgren
and
Jacob Brinton
17 Feb 2026 16:55 UTC
15
points
2
comments
5
min read
LW
link
No comments.
Back to top