Dynamic one-time delivery of critical data by small and sparse UAV swarms: a model problem for MARL scaling studies
arXiv:2512.09682v2 Announce Type: replace-cross Abstract: This work studies the application of Multi-Agent Reinforcement Learning (MARL) to decentralized control of unmanned aerial vehicles to relay a