Redundant or Necessary? A Benchmark for Detecting Redundant Steps in Agent Trajectories
DGX agentarXiv:2605.29893v1 Announce Type: new Abstract: LLM-based agents have demonstrated strong capabilities in solving complex tasks through multi-step reasoning and tool use. However, existing evaluation