English

Accelerating Approximate Thompson Sampling with Underdamped Langevin Monte Carlo

Machine Learning 2024-06-24 v3 Artificial Intelligence Machine Learning

Abstract

Approximate Thompson sampling with Langevin Monte Carlo broadens its reach from Gaussian posterior sampling to encompass more general smooth posteriors. However, it still encounters scalability issues in high-dimensional problems when demanding high accuracy. To address this, we propose an approximate Thompson sampling strategy, utilizing underdamped Langevin Monte Carlo, where the latter is the go-to workhorse for simulations of high-dimensional posteriors. Based on the standard smoothness and log-concavity conditions, we study the accelerated posterior concentration and sampling using a specific potential function. This design improves the sample complexity for realizing logarithmic regrets from O~(d)\mathcal{\tilde O}(d) to O~(d)\mathcal{\tilde O}(\sqrt{d}). The scalability and robustness of our algorithm are also empirically validated through synthetic experiments in high-dimensional bandit problems.

Keywords

Cite

@article{arxiv.2401.11665,
  title  = {Accelerating Approximate Thompson Sampling with Underdamped Langevin Monte Carlo},
  author = {Haoyang Zheng and Wei Deng and Christian Moya and Guang Lin},
  journal= {arXiv preprint arXiv:2401.11665},
  year   = {2024}
}

Comments

52 pages, 2 figures

R2 v1 2026-06-28T14:23:06.136Z