English
Related papers

Related papers: On data-based optimal stopping under stationarity …

200 papers

We study the infinite-horizon restless bandit problem with the average reward criterion, in both discrete-time and continuous-time settings. A fundamental goal is to efficiently compute policies that achieve a diminishing optimality gap as…

Machine Learning · Computer Science 2024-01-17 Yige Hong , Qiaomin Xie , Yudong Chen , Weina Wang

This note re-visits the rolling-horizon control approach to the problem of a Markov decision process (MDP) with infinite-horizon discounted expected reward criterion. Distinguished from the classical value-iteration approach, we develop an…

Optimization and Control · Mathematics 2022-06-07 Hyeong Soo Chang

An iterative learning algorithm is presented for continuous-time linear-quadratic optimal control problems where the system is externally symmetric with unknown dynamics. Both finite-horizon and infinite-horizon problems are considered. It…

Optimization and Control · Mathematics 2025-10-10 Hamed Taghavian , Florian Dorfler , Mikael Johansson

We consider a class of time-inhomogeneous optimal stopping problems and we provide sufficient conditions on the data of the problem that guarantee monotonicity of the optimal stopping boundary. In our setting, time-inhomogeneity stems not…

Optimization and Control · Mathematics 2023-01-16 Alessandro Milazzo

For optimal control problems on finite graphs in continuous time, the dynamic programming principle leads to value functions characterized by systems of nonlinear ordinary differential equations. In this paper, we consider the case of…

Optimization and Control · Mathematics 2022-12-29 Olivier Guéant

This paper considers an optimal impulse control problem of dynamical systems generated by a flow. The performance criteria are total costs over the infinite time horizon. Apart from the main performance to be minimized, there are multiple…

Optimization and Control · Mathematics 2020-10-27 Alexey Piunovskiy , Yi Zhang

In this paper, we consider algorithms with integral action for solving online optimization problems characterized by quadratic cost functions with a time-varying optimal point described by an $(n-1)$th order polynomial. Using a version of…

Optimization and Control · Mathematics 2025-09-12 Alex Xinting Wu , Ian R. Petersen , Iman Shames

In this note we propose a new approach towards solving numerically optimal stopping problems via reinforced regression based Monte Carlo algorithms. The main idea of the method is to reinforce standard linear regression algorithms in each…

Numerical Analysis · Mathematics 2019-07-02 Denis Belomestny , John Schoenmakers , Vladimir Spokoiny , Bakhyt Zharkynbay

This paper uses recent results on continuous-time finite-horizon optimal switching problems with negative switching costs to prove the existence of a saddle point in an optimal stopping (Dynkin) game. Sufficient conditions for the game's…

Optimization and Control · Mathematics 2018-06-05 Randall Martyr

The inverse linear-quadratic optimal control problem is a system identification problem whose aim is to recover the quadratic cost function and hence the closed-loop system matrices based on observations of optimal trajectories. In this…

Optimization and Control · Mathematics 2022-09-22 Han Zhang , Axel Ringh

We propose an algorithm for non-stationary kernel bandits that does not require prior knowledge of the degree of non-stationarity. The algorithm follows randomized strategies obtained by solving optimization problems that balance…

Machine Learning · Statistics 2023-02-21 Kihyuk Hong , Yuhang Li , Ambuj Tewari

Presented is an algorithm to synthesize the optimal infinite-horizon LQR feedback controller for continuous-time systems. The algorithm does not require knowledge of the system dynamics but instead uses only a finite-length sampling of…

Optimization and Control · Mathematics 2026-02-16 Sean Bowerfind , Matthew R. Kirchner , Gary Hewer

Time-sensitive machine learning benefits from Sequential Probability Ratio Test (SPRT), which provides an optimal stopping time for early classification of time series. However, in finite horizon scenarios, where input lengths are finite,…

Machine Learning · Computer Science 2025-01-31 Akinori F. Ebihara , Taiki Miyagawa , Kazuyuki Sakurai , Hitoshi Imaoka

In this work, an innovative data-driven moving horizon state estimation is proposed for model dynamic-unknown systems based on Bayesian optimization. As long as the measurement data is received, a locally linear dynamics model can be…

Systems and Control · Electrical Eng. & Systems 2023-11-14 Qing Sun , Shuai Niu , Minrui Fei

Recently, suboptimality estimates for model predictive controllers (MPC) have been derived for the case without additional stabilizing endpoint constraints or a Lyapunov function type endpoint weight. The proposed methods yield a posteriori…

Optimization and Control · Mathematics 2015-03-19 Thomas Jahn , Jürgen Pannek

The goal of a learner, in standard online learning, is to have the cumulative loss not much larger compared with the best-performing function from some fixed class. Numerous algorithms were shown to have this gap arbitrarily close to zero,…

Machine Learning · Computer Science 2013-03-04 Nina Vaits , Edward Moroshko , Koby Crammer

Sample-efficient exploration is crucial not only for discovering rewarding experiences but also for adapting to environment changes in a task-agnostic fashion. A principled treatment of the problem of optimal input synthesis for system…

Machine Learning · Computer Science 2019-10-10 Matthias Schultheis , Boris Belousov , Hany Abdulsamad , Jan Peters

This paper is concerned with devising the nonlinear optimal guidance for intercepting a stationary target with a fixed impact time. According to Pontryagin's Maximum Principle (PMP), some optimality conditions for the solutions of the…

Optimization and Control · Mathematics 2024-01-04 Kun Wang , Zheng Chen , Han Wang , Jun Li

We consider the problem of optimally stopping a general one-dimensional stochastic differential equation (SDE) with generalised drift over an infinite time horizon. First, we derive a complete characterisation of the solution to this…

Probability · Mathematics 2019-09-26 Mihail Zervos , Neofytos Rodosthenous , Pui Chan Lon , Thomas Bernhardt

This paper presents a computationally efficient algorithm for eco-driving over long prediction horizons. The eco-driving problem is formulated as a bi-level program, where the bottom level is solved offline, pre-optimizing gear as a…

Systems and Control · Electrical Eng. & Systems 2020-02-07 Ahad Hamednia , Nalin Kumar Sharma , Nikolce Murgovski , Jonas Fredriksson