Consequentialist Objectives and Catastrophe
DGX agentarXiv:2603.15017v3 Announce Type: replace Abstract: Because human preferences are too complex to codify, AIs operate with misspecified objectives. Optimizing such objectives often produces undesirable