English

Learning to Cover: Online Learning and Optimization with Irreversible Decisions

Machine Learning 2026-03-06 v3 Optimization and Control

Abstract

We define an online learning and optimization problem with discrete and irreversible decisions contributing toward a coverage target. In each period, a decision-maker selects facilities to open, receives information on the success of each one, and updates a classification model to guide future decisions. The goal is to minimize facility openings under a chance constraint reflecting the coverage target, in an asymptotic regime characterized by a large target number of facilities mm\to\infty but a finite horizon TZ+T \in \mathcal{Z}_+. We prove that, under statistical conditions, the online classifier converges to the Bayes-optimal classifier at a rate of at best O(1/n)\mathcal{O}(1/\sqrt n). Thus, we formulate our online learning and optimization problem, with a generalized learning rate r>0r>0 and a residual error 1p1-p. We derive an asymptotically optimal algorithm and an asymptotically tight lower bound. The regret grows in Θ(m1r1rT)\Theta\left(m^{\frac{1-r}{1-r^T}}\right) if p=1p=1 (perfect learning) or in Θ(max{m1r1rT,m})\Theta\left(\max\left\{m^{\frac{1-r}{1-r^T}},\sqrt{m}\right\}\right) otherwise; in particular, the regret rate is sub-linear and converges exponentially fast to its infinite-horizon limit. We extend this result to a more complicated facility location setting in a bipartite facility-customer graph with a target on customer coverage. Throughout, constructive proofs identify a policy featuring limited exploration initially and fast exploitation later on once uncertainty gets mitigated. These results uncover the benefits of limited online learning and optimization through pilot programs prior to full-fledged expansion.

Keywords

Cite

@article{arxiv.2406.14777,
  title  = {Learning to Cover: Online Learning and Optimization with Irreversible Decisions},
  author = {Alexandre Jacquillat and Michael Lingzhi Li},
  journal= {arXiv preprint arXiv:2406.14777},
  year   = {2026}
}
R2 v1 2026-06-28T17:14:09.814Z