Investigating and Alleviating Harm Amplification in LLM Interactions
DGX agentarXiv:2606.02423v1 Announce Type: new Abstract: Large language models (LLMs) can serve as helpful assistants, yet they can equally function as harm amplifiers that enable malicious users to achieve ha