English
Related papers

Related papers: Stopping problems with an unknown state

200 papers

We solve non-Markovian optimal switching problems in discrete time on an infinite horizon, when the decision maker is risk aware and the filtration is general, and establish existence and uniqueness of solutions for the associated reflected…

Optimization and Control · Mathematics 2022-08-09 Randall Martyr , John Moriarty , Magnus Perninge

Reinforcement learning in environments with many action-state pairs is challenging. At issue is the number of episodes needed to thoroughly search the policy space. Most conventional heuristics address this search problem in a stochastic…

Artificial Intelligence · Computer Science 2018-03-06 Isaac J. Sledge , Matthew S. Emigh , Jose C. Principe

We propose and analyze a continuous-time robust reinforcement learning framework for optimal stopping under ambiguity. In this framework, an agent chooses a robust exploratory stopping time motivated by two objectives: robust…

Optimization and Control · Mathematics 2026-04-17 Junyan Ye , Hoi Ying Wong , Kyunghyun Park

The study of Markov models is central to control theory and machine learning. A quantum analogue of partially observable Markov decision process was studied in (Barry, Barry, and Aaronson, Phys. Rev. A, 90, 2014). It was proved that…

Quantum Physics · Physics 2019-11-06 Christino Tamon , Weichen Xie

The goal of this paper is to analyze distributional Markov Decision Processes as a class of control problems in which the objective is to learn policies that steer the distribution of a cumulative reward toward a prescribed target law,…

Optimization and Control · Mathematics 2026-02-09 Nicole Bäuerle , Athanasios Vasileiadis

We introduce a unified probabilistic framework for solving sequential decision making problems ranging from Bayesian optimisation to contextual bandits and reinforcement learning. This is accomplished by a probabilistic model-based approach…

The problem of stopping a Brownian bridge with an unknown pinning point to maximise the expected value at the stopping time is studied. A few general properties, such as continuity and various bounds of the value function, are established.…

Probability · Mathematics 2019-01-17 Erik Ekström , Juozas Vaicenavicius

Consider the following multi-phase project management problem. Each project is divided into several phases. All projects enter the next phase at the same point chosen by the decision maker based on observations up to that point. Within each…

Statistics Theory · Mathematics 2007-06-13 Hock Peng Chan , Cheng-Der Fuh , Inchi Hu

We consider the problem of stabilization of a linear system, under state and control constraints, and subject to bounded disturbances and unknown parameters in the state matrix. First, using a simple least square solution and available…

Systems and Control · Electrical Eng. & Systems 2020-07-22 Edouard Leurent , Denis Efimov , Odalric-Ambrym Maillard

We study the problem of optimally managing an inventory with unknown demand trend. Our formulation leads to a stochastic control problem under partial observation, in which a Brownian motion with non-observable drift can be singularly…

Optimization and Control · Mathematics 2022-11-28 Salvatore Federico , Giorgio Ferrari , Neofytos Rodosthenous

This paper considers stochastic-constrained stochastic optimization where the stochastic constraint is to satisfy that the expectation of a random function is below a certain threshold. In particular, we study the setting where data samples…

Optimization and Control · Mathematics 2026-01-27 Yeongjong Kim , Dabeen Lee

Safety-critical cyber-physical systems require control strategies whose worst-case performance is robust against adversarial disturbances and modeling uncertainties. In this paper, we present a framework for approximate control and learning…

Optimization and Control · Mathematics 2023-04-04 Aditya Dave , Ioannis Faros , Nishanth Venkatesh , Andreas A. Malikopoulos

Optimal liquidation of an asset with unknown constant drift and stochastic regime-switching volatility is studied. The uncertainty about the drift is represented by an arbitrary probability distribution; the stochastic volatility is…

Mathematical Finance · Quantitative Finance 2019-01-17 Juozas Vaicenavicius

This paper derives an optimal control strategy for a simple stochastic dynamical system with constant drift and an additive control input. Motivated by the example of a physical system with an unexpected change in its dynamics, we take the…

Optimization and Control · Mathematics 2022-02-09 Daniel Gurevich , Debdipta Goswami , Charles L. Fefferman , Clarence W. Rowley

We exhibit optimal control strategies for a simple toy problem in which the underlying dynamics depend on a parameter that is initially unknown and must be learned. We consider a cost function posed over a finite time interval, in contrast…

Optimization and Control · Mathematics 2020-02-27 Charles L. Fefferman , Bernat Guillen Pegueroles , Clarence W. Rowley , Melanie Weber

We study the Markowitz portfolio selection problem with unknown drift vector in the multidimensional framework. The prior belief on the uncertain expected rate of return is modeled by an arbitrary probability law, and a Bayesian approach…

Portfolio Management · Quantitative Finance 2018-11-19 Carmine De Franco , Johann Nicolle , Huyên Pham

Scenario reduction algorithms can be an effective means to provide a tractable description of the uncertainty in optimal control problems. However, they might significantly compromise the performance of the controlled system. In this paper,…

Optimization and Control · Mathematics 2024-04-12 Francesco Cordiano , Bart De Schutter

We analyze an optimal stopping problem with a series of inequality-type and equality-type expectation constraints in a general non-Markovian framework. We show that the optimal stopping problem with expectation constraints (OSEC) in an…

Optimization and Control · Mathematics 2023-02-10 Erhan Bayraktar , Song Yao

A central problem in business concerns the optimal allocation of limited resources to a set of available tasks, where the payoff of these tasks is inherently uncertain. In credit card fraud detection, for instance, a bank can only assign a…

Machine Learning · Computer Science 2022-02-10 Toon Vanderschueren , Bart Baesens , Tim Verdonck , Wouter Verbeke

In this paper, we study the remote estimation problem of a Markov process over a channel with a cost. We formulate this problem as an infinite horizon optimization problem with two players, i.e., a sensor and a monitor, that have distinct…

Systems and Control · Electrical Eng. & Systems 2024-02-01 Edoardo David Santi , Touraj Soleymani , Deniz Gunduz