Trust No Tool: Evaluating and Defending LLM Agents under Untrusted Tool Feedback
DGX agentarXiv:2605.17453v1 Announce Type: cross Abstract: Tool-using LLM agents increasingly rely on external tools to make consequential decisions, yet most existing agent-security benchmarks and defenses im