Gate AI: LLM Security Benchmark Evaluation Methodology and Results
DGX agentarXiv:2606.02959v1 Announce Type: new Abstract: Published evaluations of prompt-injection and jailbreak detectors for Large Language Models often suffer from two systematic weaknesses: per-dataset thr