3 papers
cs.LG2026
Balancing Expressivity and Learnability in Quantum Kernel Bandit Optimization
Yuqi Huang, Vincent Y. F. Tan, Sharu Theresa Jose
We investigate Gaussian process (GP) bandit optimization with quantum kernels, assuming the mean reward function lies in the reproducing kernel Hilbert space (RKHS) induced by the…
cs.LG2026
Almost Asymptotically Optimal Active Clustering Through Pairwise Observations
Rachel S. Y. Teo, P. N. Karthik, Ramya Korlakai Vinayak +1
We propose a new analysis framework for clustering items into an unknown number of distinct groups using noisy and actively collected responses. At each time step, an agent…
cs.LG2026
Quantum-Enhanced Neural Contextual Bandit Algorithms
Yuqi Huang, Vincent Y. F Tan, Sharu Theresa Jose
Stochastic contextual bandits are fundamental for sequential decision-making but pose significant challenges for existing neural network-based algorithms, particularly when scaling…