Safety

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍 If anyone builds it, everyone thrives. Over the past decade, a lot of important work on AI alignment has focus

DGX agentx-post
safetyyohei-nakajima--x

tl;dr - let's not just remove negative behavior from models, but also add positive ones 👍 If anyone builds it, everyone thrives. Over the past decade, a lot of important work on AI alignment has focused on avoiding harm. But freedom from harm isn't the same as freedom to flourish. In this paper, we introduce 'Positive Alignment'. A positively aligned agent is one that…

Source: Yohei Nakajima (X) | 2026-05-12

Loading related sources…