Improved Algorithms for Nash Welfare in Linear Bandits
DGX agentarXiv:2601.22969v2 Announce Type: replace Abstract: Nash regret has recently emerged as a principled fairness-aware performance metric for stochastic multi-armed bandits, motivated by the Nash Social