Bridging the Evaluation Gap: Standardized Benchmarks for Multi-Objective Search
DGX agentarXiv:2603.24084v2 Announce Type: replace Abstract: Empirical evaluation in multi-objective search (MOS) has historically suffered from fragmentation, relying on heterogeneous problem instances with i