English
Related papers

Related papers: The Spend-It-All Region and Small Time Results for…

200 papers

We consider a singular stochastic control problem, which is called the Monotone Follower Stochastic Control Problem and give sufficient conditions for the existence and uniqueness of a local-time type optimal control. To establish this…

Optimization and Control · Mathematics 2007-05-23 Erhan Bayraktar , Masahiko Egami

We study a generalization of the multi-armed bandit problem with multiple plays where there is a cost associated with pulling each arm and the agent has a budget at each time that dictates how much she can expect to spend. We derive an…

Machine Learning · Statistics 2019-09-13 Alexander Luedtke , Emilie Kaufmann , Antoine Chambaz

The airplane refueling problem is a nonlinear combinatorial optimization problem, and its equivalent problem the $n$-vehicle exploration problem is proved to be NP-complete (arXiv:2304.03965v1, The $n$-vehicle exploration problem is…

Computational Complexity · Computer Science 2023-05-23 Jinchuan Cui , Xiaoya Li

We consider the contextual bandit problem on general action and context spaces, where the learner's rewards depend on their selected actions and an observable context. This generalizes the standard multi-armed bandit to the case where side…

Machine Learning · Statistics 2023-01-03 Moise Blanchard , Steve Hanneke , Patrick Jaillet

Budgeted uncertainty sets have been established as a major influence on uncertainty modeling for robust optimization problems. A drawback of such sets is that the budget constraint only restricts the global amount of cost increase that can…

Optimization and Control · Mathematics 2020-08-28 Marc Goerigk , Stefan Lendl

A class of optimal control problems governed by semilinear parabolic equations with mixed constraints and a box constraint for control variable is considered. We show that if the separation condition is satisfied, then both optimality…

Optimization and Control · Mathematics 2023-09-06 Huynh Khanh , Bui Trong Kien

We solve the problem of optimal stopping of a Brownian motion subject to the constraint that the stopping time's distribution is a given measure consisting of finitely-many atoms. In particular, we show that this problem can be converted to…

Optimization and Control · Mathematics 2017-07-07 Erhan Bayraktar , Christopher W. Miller

The basic problem of optimal transportation consists in minimizing the expected costs $\mathbb {E}[c(X_1,X_2)]$ by varying the joint distribution $(X_1,X_2)$ where the marginal distributions of the random variables $X_1$ and $X_2$ are…

Probability · Mathematics 2016-08-14 Mathias Beiglböck , Nicolas Juillet

We consider the best-choice problem for independent (not necessarily iid) observations $X_1, \cdots, X_n$ with the aim of selecting the sample minimum. We show that in this full generality the monotone case of optimal stopping holds and the…

Probability · Mathematics 2021-10-13 Alexander Gnedin , Patryk Kozieł , Małgorzata Sulkowska

We consider the problem of stochastic $K$-armed dueling bandit in the contextual setting, where at each round the learner is presented with a context set of $K$ items, each represented by a $d$-dimensional feature vector, and the goal of…

Machine Learning · Computer Science 2021-05-11 Aadirupa Saha , Aditya Gopalan

In this article, we consider a species whose population density solves the steady diffusive logistic equation in a heterogeneous environment modeled with the help of a spatially non constant coefficient standing for a resources…

Analysis of PDEs · Mathematics 2019-07-30 Idriss Mazari , Grégoire Nadin , Yannick Privat

We present a Defense/Attack resource allocation model, where Defender has some number of ``locks" to protect $n$ vulnerable boxes (sites), and Attacker is trying to destroy these boxes, having $m$ ``bombs" that can be placed into the boxes.…

Computer Science and Game Theory · Computer Science 2023-07-06 Isaac M. Sonin

We investigate an expected utility maximization problem under model uncertainty in a one-period financial market. We capture model uncertainty by replacing the baseline model $\mathbb{P}$ with an adverse choice from a Wasserstein ball of…

Optimization and Control · Mathematics 2024-01-17 Laurence Carassus , Johannes Wiesel

Consider the set of probability measures with given marginal distributions on the product of two complete, separable metric spaces, seen as a correspondence when the marginal distributions vary. In problems of optimal transport, continuity…

Risk Management · Quantitative Finance 2020-09-29 Mario Ghossoub , David Saunders

We study a sequential decision problem where the learner faces a sequence of $K$-armed bandit tasks. The task boundaries might be known (the bandit meta-learning setting), or unknown (the non-stationary bandit setting). For a given integer…

The Stackelberg game model, where a leader commits to a strategy and the follower best responds, has found widespread application, particularly to security problems. In the security setting, the goal is for the leader to compute an optimal…

Computer Science and Game Theory · Computer Science 2022-09-19 Sai Mali Ananthanarayanan , Christian Kroer

We consider the infinite-horizon average-reward restless bandit problem. We propose a novel \emph{two-set policy} that maintains two dynamic subsets of arms: one subset of arms has a nearly optimal state distribution and takes actions…

Machine Learning · Computer Science 2024-10-18 Yige Hong , Qiaomin Xie , Yudong Chen , Weina Wang

In this paper, we investigate an optimal design problem motivated by some issues arising in population dynamics. In a nutshell, we aim at determining the optimal shape of a region occupied by resources for maximizing the survival ability of…

Analysis of PDEs · Mathematics 2017-09-08 Fabien Caubet , Thibaut Deheuvels , Yannick Privat

We consider the problem of designing contextual bandit algorithms in the ``cross-learning'' setting of Balseiro et al., where the learner observes the loss for the action they play in all possible contexts, not just the context of the…

Machine Learning · Computer Science 2024-01-04 Jon Schneider , Julian Zimmert

In the classical many normal means with different variances, we consider the situation when the observer is allowed to allocate the available measurement budget over the coordinates of the parameter of interest. The benchmark is the minimax…

Statistics Theory · Mathematics 2019-04-01 Eduard Belitser