Claudini: Autoresearch Discovers State-of-the-Art Adversarial Attack Algorithms for LLMs
DGX agentarXiv:2603.24511v2 Announce Type: replace-cross Abstract: We show that AI agents are capable of discovering novel algorithms for adversarial attacks against LLMs, advancing the state of the art on whi