MemSearcher: Training LLMs to Reason, Search and Manage Memory via End-to-End Reinforcement Learning
DGX agentarXiv:2511.02805v2 Announce Type: replace-cross Abstract: LLM-based search agents often concatenate the full interaction history into the context, producing long and noisy inputs, and increasing compu