Alignment as Institutional Design: From Behavioral Correction to Transaction Structure in Intelligent Systems
DGX agentarXiv:2604.13079v1 Announce Type: cross Abstract: Current AI alignment paradigms rely on behavioral correction: external supervisors (e.g., RLHF) observe outputs, judge against preferences, and adjust