Offline-to-Online Learning in Linear Bandits
DGX agentarXiv:2606.04305v1 Announce Type: new Abstract: We study online learning with an additional offline dataset in the stochastic linear bandit setting. Although this problem arises frequently in practice