Moral Safety in LLMs: Exposing Performative Compliance with Puzzled Cues
DGX agentarXiv:2606.31644v1 Announce Type: new Abstract: As large language models take on morally consequential roles in healthcare, legal, and hiring contexts, we need to examine whether their ethical behavio