Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games
DGX agentarXiv:2510.24515v2 Announce Type: replace Abstract: The Team Orienteering Problem (TOP) generalizes many real-world multi-agent scheduling and routing tasks that occur in autonomous mobility, aerial l