Best of many worlds: Robust model selection for online supervised learning
Abstract
We introduce algorithms for online, full-information prediction that are competitive with contextual tree experts of unknown complexity, in both probabilistic and adversarial settings. We show that by incorporating a probabilistic framework of structural risk minimization into existing adaptive algorithms, we can robustly learn not only the presence of stochastic structure when it exists (leading to constant as opposed to regret), but also the correct model order. We thus obtain regret bounds that are competitive with the regret of an optimal algorithm that possesses strong side information about both the complexity of the optimal contextual tree expert and whether the process generating the data is stochastic or adversarial. These are the first constructive guarantees on simultaneous adaptivity to the model and the presence of stochasticity.
Cite
@article{arxiv.1805.08562,
title = {Best of many worlds: Robust model selection for online supervised learning},
author = {Vidya Muthukumar and Mitas Ray and Anant Sahai and Peter L. Bartlett},
journal= {arXiv preprint arXiv:1805.08562},
year = {2018}
}
Comments
33 pages, 5 figures