English
Related papers

Related papers: The Spend-It-All Region and Small Time Results for…

200 papers

In this paper, we present approximation algorithms for combinatorial optimization problems under probabilistic constraints. Specifically, we focus on stochastic variants of two important combinatorial optimization problems: the k-center…

Data Structures and Algorithms · Computer Science 2008-09-03 Shipra Agrawal , Amin Saberi , Yinyu Ye

The challenge of identifying the best feasible arm within a fixed budget has attracted considerable interest in recent years. However, a notable gap remains in the literature: the exact exponential rate at which the error probability…

Machine Learning · Computer Science 2025-06-04 Jie Bian , Vincent Y. F. Tan

This paper is concerned with the existence and regularity of mininizers as well as of corresponding multipliers to an optimal control problem governed by semilinear elliptic equations, in which mixed pointwise control-state constraints are…

Optimization and Control · Mathematics 2023-11-28 Vu Huu Nhu , Nguyen Quoc Tuan , Nguyen Bang Giang , Nguyen Thi Thu Huong

We consider a budget-constrained bandit problem where each arm pull incurs a random cost, and yields a random reward in return. The objective is to maximize the total expected reward under a budget constraint on the total cost. The model is…

Machine Learning · Computer Science 2020-03-03 Semih Cayci , Atilla Eryilmaz , R. Srikant

Boolean satisfiability [1] (k-SAT) is one of the most studied optimization problems, as an efficient (that is, polynomial-time) solution to k-SAT (for $k\geq 3$) implies efficient solutions to a large number of hard optimization problems…

Computational Complexity · Computer Science 2012-08-03 Maria Ercsey-Ravasz , Zoltan Toroczkai

In this paper, we study the optimal stopping problem in the so-called exploratory framework, in which the agent takes actions randomly conditioning on current state and an entropy-regularized term is added to the reward functional. Such a…

Optimization and Control · Mathematics 2023-09-04 Yuchao Dong

We address prehensile pushing, the problem of manipulating a grasped object by pushing against the environment. Our solution is an efficient nonlinear trajectory optimization problem relaxed from an exact mixed integer non-linear trajectory…

Robotics · Computer Science 2025-03-19 Patrizio Perugini , Jens Lundell , Katharina Friedl , Danica Kragic

In a multi-armed bandit problem, an online algorithm chooses from a set of strategies in a sequence of trials so as to maximize the total payoff of the chosen strategies. While the performance of bandit algorithms with a small finite…

Data Structures and Algorithms · Computer Science 2008-09-30 Robert Kleinberg , Aleksandrs Slivkins , Eli Upfal

Boltzmann exploration is a classic strategy for sequential decision-making under uncertainty, and is one of the most standard tools in Reinforcement Learning (RL). Despite its widespread use, there is virtually no theoretical understanding…

Machine Learning · Computer Science 2017-11-08 Nicolò Cesa-Bianchi , Claudio Gentile , Gábor Lugosi , Gergely Neu

We present a methodology for obtaining explicit solutions to infinite time horizon optimal stopping problems involving general, one-dimensional, It\^o diffusions, payoff functions that need not be smooth and state-dependent discounting.…

Computational Finance · Quantitative Finance 2012-10-10 Timothy C. Johnson

In this article, we give an in-depth analysis of the problem of optimising the total population size for a standard logistic-diffusive model. This optimisation problem stems from the study of spatial ecology and amounts to the following…

Analysis of PDEs · Mathematics 2021-05-24 Idriss Mazari , Grégoire Nadin , Yannick Privat

We study a problem when a solution to optimal stopping problem for one-dimensional diffusion will generate by threshold strategy. Namely, we give necessary and sufficient conditions under which an optimal stopping time can be specified as…

Probability · Mathematics 2013-06-20 V. I. Arkin , A. D. Slastnikov

We consider a multi-armed bandit problem with $M$ latent contexts, where an agent interacts with the environment for an episode of $H$ time steps. Depending on the length of the episode, the learner may not be able to estimate accurately…

Machine Learning · Computer Science 2022-10-10 Jeongyeol Kwon , Yonathan Efroni , Constantine Caramanis , Shie Mannor

We consider the problem of estimating the total probability of all symbols that appear with a given frequency in a string of i.i.d. random variables with unknown distribution. We focus on the regime in which the block length is large yet no…

Information Theory · Computer Science 2016-11-15 Aaron B. Wagner , Pramod Viswanath , Sanjeev R. Kulkarni

We consider the best arm identification problem, where the goal is to identify the arm with the highest mean reward from a set of $K$ arms under a limited sampling budget. This problem models many practical scenarios such as A/B testing. We…

Machine Learning · Statistics 2026-05-05 Junpei Komiyama , Kyoungseok Jang , Junya Honda

We consider the spatially inhomogeneous Landau equation with soft potentials, including the case of Coulomb interactions. First, we establish the existence of solutions for a short time, assuming the initial data is in a fourth-order…

Analysis of PDEs · Mathematics 2018-05-30 Christopher Henderson , Stanley Snelson , Andrei Tarfulea

The submodular maximization problem is widely applicable in many engineering problems where objectives exhibit diminishing returns. While this problem is known to be NP-hard for certain subclasses of objective functions, there is a greedy…

Distributed, Parallel, and Cluster Computing · Computer Science 2020-07-01 Haoyuan Sun , David Grimsman , Jason R Marden

The regularity of solutions to the Boltzmann equation is a fundamental problem in the kinetic theory. In this paper, the case with angular cut-off is investigated. It is shown that the macroscopic parts of solutions to the Boltzmann…

Analysis of PDEs · Mathematics 2016-01-07 Feimin Huang , Yong Wang

The $k$-of-$n$ testing problem involves performing $n$ independent tests sequentially, in order to determine whether/not at least $k$ tests pass. The objective is to minimize the expected cost of testing. This is a fundamental and…

Data Structures and Algorithms · Computer Science 2026-03-26 Rayen Tan , Viswanath Nagarajan

We prove general theorems for isoperimetric problems on lattices of the form ${\mathbb{Z}}^{k} \times {\mathbb{N}}^{d}$ which state that the perimeter of the optimal set is a monotonically increasing function of the volume under certain…

Combinatorics · Mathematics 2013-09-10 Emmanuel Tsukerman