SEVRA-BENCH: Social Engineering of Vulnerabilities in Review Agents
DGX agentarXiv:2606.13757v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed in automated code-review systems, where their approvals can determine which code is mer