English
Related papers

Related papers: Grab It Before It's Gone: Testing Uncertain Reward…

200 papers

Decision-making under uncertainty is a fundamental problem encountered frequently and can be formulated as a stochastic multi-armed bandit problem. In the problem, the learner interacts with an environment by choosing an action at each…

Machine Learning · Statistics 2024-05-24 Jonathan Gornet , Bruno Sinopoli

Stochastic two-player games model systems with an environment that is both adversarial and stochastic. In this paper, we study the expected value of bounded quantitative prefix-independent objectives in the context of stochastic games. We…

Computer Science and Game Theory · Computer Science 2025-08-01 Laurent Doyen , Pranshu Gaba , Shibashis Guha

We consider a sequential decision-making problem where an agent can take one action at a time and each action has a stochastic temporal extent, i.e., a new action cannot be taken until the previous one is finished. Upon completion, the…

Machine Learning · Computer Science 2020-03-26 P Sharoff , Nishant A. Mehta , Ravi Ganti

This paper derives several formulae for the probability that a Wiener process, which has a stochastic drift and random variance, crosses a one-sided stochastic boundary within a finite time interval. A non-explicit formula is first obtained…

Probability · Mathematics 2024-10-04 Yoann Potiron

We study decision timing problems on finite horizon with Poissonian information arrivals. In our model, a decision maker wishes to optimally time her action in order to maximize her expected reward. The reward depends on an unobservable…

Optimization and Control · Mathematics 2012-05-07 Michael Ludkovski , Semih Sezer

The analysis of dynamical systems is a fundamental tool in the natural sciences and engineering. It is used to understand the evolution of systems as large as entire galaxies and as small as individual molecules. With predefined conditions…

Machine Learning · Statistics 2024-12-19 Ludwig Winkler

This work deals with the one-dimensional Stefan problem with a general time-dependent boundary condition at the fixed boundary. Stochastic solutions are obtained using discrete random walks, and the results are compared with analytic…

Analysis of PDEs · Mathematics 2023-02-06 M. Ogren

We obtain the first probabilistic proof of continuous differentiability of time-dependent optimal boundaries in optimal stopping problems. The underlying stochastic dynamics is a one-dimensional, time-inhomogeneous diffusion. The gain…

Probability · Mathematics 2024-05-28 Tiziano De Angelis , Damien Lamberton

We consider both discrete and continuous "uncertain horizon" deterministic control processes, for which the termination time is a random variable. We examine the dynamic programming equations for the value function of such processes,…

Optimization and Control · Mathematics 2016-01-06 June Andrews , Alexander Vladimirsky

The inverse first-passage time problem determines a boundary such that the first-passage time of a Wiener process to this boundary has a given distribution. An approximation which is based on the starting value of the boundary to a smooth…

Probability · Mathematics 2023-09-06 Yoann Potiron

Stochastic parabolic equations are widely used to model many random phenomena in natural sciences, such as the temperature distribution in a noisy medium, the dynamics of a chemical reaction in a noisy environment, or the evolution of the…

Analysis of PDEs · Mathematics 2023-09-21 Zhonghua Liao , Qi Lü

We propose a new model for formalizing reward collection problems on graphs with dynamically generated rewards which may appear and disappear based on a stochastic model. The *robot routing problem* is modeled as a graph whose nodes are…

Systems and Control · Computer Science 2017-07-18 Rayna Dimitrova , Ivan Gavran , Rupak Majumdar , Vinayak S. Prabhu , Sadegh Esmaeil Zadeh Soudjani

We consider Thompson sampling for linear bandit problems with finitely many independent arms, where rewards are sampled from normal distributions that are linearly dependent on unknown parameter vectors and with unknown variance.…

Machine Learning · Computer Science 2023-03-07 Björn Lindenberg , Karl-Olof Lindahl

We analyze an irreversible investment decision for a project which yields a flow of future operating profits given by a geometric Brownian motion with unknown drift. In contrast to similar optimal stopping problems with incomplete…

Optimization and Control · Mathematics 2025-02-19 Fabian Gierens , Berenice Anne Neumann

Many sequential decision-making problems in communication networks can be modeled as contextual bandit problems, which are natural extensions of the well-known multi-armed bandit problem. In contextual bandit problems, at each time, an…

Machine Learning · Computer Science 2016-05-10 Pranav Sakulkar , Bhaskar Krishnamachari

Given a Wiener process with unknown and unobservable drift, we try to estimate this drift as effectively but also as quickly as possible, in the presence of a quadratic penalty for the estimation error and of a fixed, positive cost per unit…

Statistics Theory · Mathematics 2019-05-24 Erik Ekström , Ioannis Karatzas , Juozas Vaicenavicius

When optimizing problems with uncertain parameter values in a linear objective, decision-focused learning enables end-to-end learning of these values. We are interested in a stochastic scheduling problem, in which processing times are…

Machine Learning · Computer Science 2024-08-16 Kim van den Houten , David M. J. Tax , Esteban Freydell , Mathijs de Weerdt

Free boundary problems appear naturally in numerous areas of mathematics, science and engineering. These problems present a great computational challenge because they necessitate numerical methods that can yield an accurate approximation of…

Numerical Analysis · Mathematics 2020-12-29 Sifan Wang , Paris Perdikaris

Although evidence integration to the boundary model has successfully explained a wide range of behavioral and neural data in decision making under uncertainty, how animals learn and optimize the boundary remains unresolved. Here, we propose…

Neural and Evolutionary Computing · Computer Science 2024-08-13 Jamal Esmaily , Rani Moran , Yasser Roudi , Bahador Bahrami

We consider a general online stochastic optimization problem with multiple budget constraints over a horizon of finite time periods. In each time period, a reward function and multiple cost functions are revealed, and the decision maker…

Machine Learning · Computer Science 2022-07-26 Jiashuo Jiang , Xiaocheng Li , Jiawei Zhang
‹ Prev 1 2 3 10 Next ›