Learning with Multiple Correct Answers -- Regret Bounds under Different Feedback Models
arXiv:2602.09402v2 Announce Type: replace Abstract: We study the problem of learning with multiple correct answers, where each instance admits a set of valid labels. We primarily focus on the online s