中文
相关论文

相关论文: The Spend-It-All Region and Small Time Results for…

200 篇论文

We consider a singular stochastic control problem, which is called the Monotone Follower Stochastic Control Problem and give sufficient conditions for the existence and uniqueness of a local-time type optimal control. To establish this…

最优化与控制 · 数学 2007-05-23 Erhan Bayraktar , Masahiko Egami

We study a generalization of the multi-armed bandit problem with multiple plays where there is a cost associated with pulling each arm and the agent has a budget at each time that dictates how much she can expect to spend. We derive an…

机器学习 · 统计学 2019-09-13 Alexander Luedtke , Emilie Kaufmann , Antoine Chambaz

The airplane refueling problem is a nonlinear combinatorial optimization problem, and its equivalent problem the $n$-vehicle exploration problem is proved to be NP-complete (arXiv:2304.03965v1, The $n$-vehicle exploration problem is…

计算复杂性 · 计算机科学 2023-05-23 Jinchuan Cui , Xiaoya Li

We consider the contextual bandit problem on general action and context spaces, where the learner's rewards depend on their selected actions and an observable context. This generalizes the standard multi-armed bandit to the case where side…

机器学习 · 统计学 2023-01-03 Moise Blanchard , Steve Hanneke , Patrick Jaillet

Budgeted uncertainty sets have been established as a major influence on uncertainty modeling for robust optimization problems. A drawback of such sets is that the budget constraint only restricts the global amount of cost increase that can…

最优化与控制 · 数学 2020-08-28 Marc Goerigk , Stefan Lendl

A class of optimal control problems governed by semilinear parabolic equations with mixed constraints and a box constraint for control variable is considered. We show that if the separation condition is satisfied, then both optimality…

最优化与控制 · 数学 2023-09-06 Huynh Khanh , Bui Trong Kien

We solve the problem of optimal stopping of a Brownian motion subject to the constraint that the stopping time's distribution is a given measure consisting of finitely-many atoms. In particular, we show that this problem can be converted to…

最优化与控制 · 数学 2017-07-07 Erhan Bayraktar , Christopher W. Miller

The basic problem of optimal transportation consists in minimizing the expected costs $\mathbb {E}[c(X_1,X_2)]$ by varying the joint distribution $(X_1,X_2)$ where the marginal distributions of the random variables $X_1$ and $X_2$ are…

概率论 · 数学 2016-08-14 Mathias Beiglböck , Nicolas Juillet

We consider the best-choice problem for independent (not necessarily iid) observations $X_1, \cdots, X_n$ with the aim of selecting the sample minimum. We show that in this full generality the monotone case of optimal stopping holds and the…

概率论 · 数学 2021-10-13 Alexander Gnedin , Patryk Kozieł , Małgorzata Sulkowska

We consider the problem of stochastic $K$-armed dueling bandit in the contextual setting, where at each round the learner is presented with a context set of $K$ items, each represented by a $d$-dimensional feature vector, and the goal of…

机器学习 · 计算机科学 2021-05-11 Aadirupa Saha , Aditya Gopalan

In this article, we consider a species whose population density solves the steady diffusive logistic equation in a heterogeneous environment modeled with the help of a spatially non constant coefficient standing for a resources…

偏微分方程分析 · 数学 2019-07-30 Idriss Mazari , Grégoire Nadin , Yannick Privat

We present a Defense/Attack resource allocation model, where Defender has some number of ``locks" to protect $n$ vulnerable boxes (sites), and Attacker is trying to destroy these boxes, having $m$ ``bombs" that can be placed into the boxes.…

计算机科学与博弈论 · 计算机科学 2023-07-06 Isaac M. Sonin

We investigate an expected utility maximization problem under model uncertainty in a one-period financial market. We capture model uncertainty by replacing the baseline model $\mathbb{P}$ with an adverse choice from a Wasserstein ball of…

最优化与控制 · 数学 2024-01-17 Laurence Carassus , Johannes Wiesel

Consider the set of probability measures with given marginal distributions on the product of two complete, separable metric spaces, seen as a correspondence when the marginal distributions vary. In problems of optimal transport, continuity…

风险管理 · 定量金融 2020-09-29 Mario Ghossoub , David Saunders

We study a sequential decision problem where the learner faces a sequence of $K$-armed bandit tasks. The task boundaries might be known (the bandit meta-learning setting), or unknown (the non-stationary bandit setting). For a given integer…

The Stackelberg game model, where a leader commits to a strategy and the follower best responds, has found widespread application, particularly to security problems. In the security setting, the goal is for the leader to compute an optimal…

计算机科学与博弈论 · 计算机科学 2022-09-19 Sai Mali Ananthanarayanan , Christian Kroer

We consider the infinite-horizon average-reward restless bandit problem. We propose a novel \emph{two-set policy} that maintains two dynamic subsets of arms: one subset of arms has a nearly optimal state distribution and takes actions…

机器学习 · 计算机科学 2024-10-18 Yige Hong , Qiaomin Xie , Yudong Chen , Weina Wang

In this paper, we investigate an optimal design problem motivated by some issues arising in population dynamics. In a nutshell, we aim at determining the optimal shape of a region occupied by resources for maximizing the survival ability of…

偏微分方程分析 · 数学 2017-09-08 Fabien Caubet , Thibaut Deheuvels , Yannick Privat

We consider the problem of designing contextual bandit algorithms in the ``cross-learning'' setting of Balseiro et al., where the learner observes the loss for the action they play in all possible contexts, not just the context of the…

机器学习 · 计算机科学 2024-01-04 Jon Schneider , Julian Zimmert

In the classical many normal means with different variances, we consider the situation when the observer is allowed to allocate the available measurement budget over the coordinates of the parameter of interest. The benchmark is the minimax…

统计理论 · 数学 2019-04-01 Eduard Belitser