This is an interesting direction, and I’m trying to understand the core problem it addresses. In general, don’t we want agents to be more capable and smarter than us, rather than constrained to something closer to our own level? My concern is that capability and alignment may not naturally track each other, a smarter agent could also end up being less aligned than we expect. So I’m curious whether the main goal here is productivity, safer personalization, or something more directly about alignment.
This is an interesting direction, and I’m trying to understand the core problem it addresses. In general, don’t we want agents to be more capable and smarter than us, rather than constrained to something closer to our own level? My concern is that capability and alignment may not naturally track each other, a smarter agent could also end up being less aligned than we expect. So I’m curious whether the main goal here is productivity, safer personalization, or something more directly about alignment.