Balancing Expressivity and Learnability in Quantum Kernel Bandit Optimization
arXiv:2607.01080v1 Announce Type: new Abstract: We investigate Gaussian process (GP) bandit optimization with quantum kernels, assuming the mean reward function lies in the reproducing kernel Hilbert