CyBiasBench: Benchmarking Bias in LLM Agents for Cyber-Attack Scenarios
DGX agentarXiv:2605.07830v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents in offensive cybersecurity. In this paper, we reveal an interesting phenom