IAPO: Input Attribution-Aware Policy Optimization for Tool Use in Small Multimodal Agents
DGX agentarXiv:2606.11652v1 Announce Type: new Abstract: This paper investigates reinforcement learning (RL) methods for improving tool-calling capabilities in multimodal small language model (SLM) agents. Whi