English
Related papers

Related papers: Stabilizing Value Iteration with and without Appro…

200 papers

For a class of quasi-variational inequalities (QVIs) of obstacle-type the stability of its solution set and associated optimal control problems are considered. These optimal control problems are non-standard in the sense that they involve…

Optimization and Control · Mathematics 2020-08-25 Amal Alphonse , Michael Hintermüller , Carlos N. Rautenberg

For a general entropy-regularized stochastic control problem on an infinite horizon, we prove that a policy iteration algorithm (PIA) converges to an optimal relaxed control. Contrary to the standard stochastic control literature, classical…

Optimization and Control · Mathematics 2026-05-14 Yu-Jui Huang , Zhenhua Wang , Zhou Zhou

A method is presented to analyze the stability of feedback systems with neural network controllers. Two stability theorems are given to prove asymptotic stability and to compute an ellipsoidal inner-approximation to the region of attraction…

Systems and Control · Electrical Eng. & Systems 2021-01-28 He Yin , Peter Seiler , Murat Arcak

We study the problem of imitating an expert demonstrator in a discrete-time, continuous state-and-action control system. We show that, even if the dynamics satisfy a control-theoretic property called exponential stability (i.e. the effects…

Machine Learning · Computer Science 2025-07-29 Max Simchowitz , Daniel Pfrommer , Ali Jadbabaie

We propose a stochastic MPC scheme using an optimization over the initial state for the predicted trajectory. Considering linear discrete-time systems under unbounded additive stochastic disturbances subject to chance constraints, we use…

Systems and Control · Electrical Eng. & Systems 2022-07-19 Henning Schlüter , Frank Allgöwer

We tackle the issue of finding a good policy when the number of policy updates is limited. This is done by approximating the expected policy reward as a sequence of concave lower bounds which can be efficiently maximized, drastically…

Artificial Intelligence · Computer Science 2016-12-30 Nicolas Le Roux

We present a method for determining optimal modes of operation for autonomously oscillating systems with uncertain parameters. In a typical application of the method, a nonlinear dynamical system is optimized with respect to an economic…

Dynamical Systems · Mathematics 2013-08-20 Darya Kastsian , Martin Mönnigmann

Predictive safety filters enable the integration of potentially unsafe learning-based control approaches and humans into safety-critical systems. In addition to simple constraint satisfaction, many control problems involve additional…

Systems and Control · Electrical Eng. & Systems 2024-09-19 Elias Milios , Kim Peter Wabersich , Felix Berkel , Lukas Schwenkel

We study the safe reinforcement learning problem with nonlinear function approximation, where policy optimization is formulated as a constrained optimization problem with both the objective and the constraint being nonconvex functions. For…

Machine Learning · Computer Science 2019-10-29 Ming Yu , Zhuoran Yang , Mladen Kolar , Zhaoran Wang

We consider a scenario where a system experiences a disruption, and the states (representing health values) of its components continue to reduce over time, unless they are acted upon by a controller. Given this dynamical setting, we…

Systems and Control · Computer Science 2020-04-09 Hemant Gehlot , Shreyas Sundaram , Satish V. Ukkusuri

The Variation Evolving Method (VEM) that originates from the continuous-time dynamics stability theory seeks the optimal solutions with variation evolution principle. After establishing the first and the second evolution equations within…

Systems and Control · Computer Science 2025-01-28 Sheng Zhang , Fei Liao , Kai-Feng He

Value Iteration (VI) is foundational to the theory and practice of modern reinforcement learning, and it is known to converge at a $\mathcal{O}(\gamma^k)$-rate, where $\gamma$ is the discount factor. Surprisingly, however, the optimal rate…

Machine Learning · Computer Science 2023-10-31 Jongmin Lee , Ernest K. Ryu

In a classical optimal stopping problem the aim is to maximize the expected value of a functional of a diffusion evaluated at a stopping time. This note considers optimal stopping problems beyond this paradigm. We study problems in which…

Probability · Mathematics 2017-08-04 Vicky Henderson , David Hobson , Matthew Zeng

This paper demonstrates many immediate connections between adaptive control and optimization methods commonly employed in machine learning. Starting from common output error formulations, similarities in update law modifications are…

Optimization and Control · Mathematics 2020-04-17 Joseph E. Gaudio , Travis E. Gibson , Anuradha M. Annaswamy , Michael A. Bolender , Eugene Lavretsky

In this paper, we study the portfolio optimization problem with general utility functions and when the return and volatility of underlying asset are slowly varying. An asymptotic optimal strategy is provided within a specific class of…

Mathematical Finance · Quantitative Finance 2016-11-08 Jean-Pierre Fouque , Ruimeng Hu

This note is concerned with the study of the initial boundary value problem for systems of conservation laws from the point of view of control theory, where the initial data is fixed and the boundary data are regarded as control functions.…

Analysis of PDEs · Mathematics 2007-05-23 F. Ancona , A. Bressan , G. M. Coclite

Reinforcement learning methods often produce brittle policies -- policies that perform well during training, but generalize poorly beyond their direct training experience, thus becoming unstable under small disturbances. To address this…

Robotics · Computer Science 2023-02-14 Sergey Pankov

Safety is the priority concern when applying reinforcement learning (RL) algorithms to real-world control problems. While policy iteration provides a fundamental algorithm for standard RL, an analogous theoretical algorithm for safe RL…

Machine Learning · Computer Science 2025-03-14 Yujie Yang , Zhilong Zheng , Shengbo Eben Li , Wei Xu , Jingjing Liu , Xianyuan Zhan , Ya-Qin Zhang

Optimal control problems are inherently hard to solve as the optimization must be performed simultaneously with updating the underlying system. Starting from an initial guess, Howard's policy improvement algorithm separates the step of…

Optimization and Control · Mathematics 2020-05-25 B. Kerimkulov , D. Šiška , Ł. Szpruch

Motivated by the applications, a class of optimal control problems is investigated, where the goal is to influence the behavior of a given population through another controlled one interacting with the first. Diffusive terms accounting for…

Optimization and Control · Mathematics 2023-03-10 Stefano Almi , Marco Morandotti , Francesco Solombrino