arxivcs.LGcs.IT2026-07-01
Balancing Expressivity and Learnability in Quantum Kernel Bandit Optimization
Yuqi Huang, Vincent Y. F. Tan, Sharu Theresa Jose
We investigate Gaussian process (GP) bandit optimization with quantum kernels, assuming the mean reward function lies in the reproducing kernel Hilbert space (RKHS) induced by the quantum kernel. This setting is motivated by NISQ-era tasks such as quantum control, state preparati…