GroupMemBench: Benchmarking LLM Agent Memory in Multi-Party Conversations
DGX agentarXiv:2605.14498v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly serve as personal assistants and workplace collaborators, where their utility depends on memory systems t