English
Related papers

Related papers: Convergence Analysis of Policy Iteration

200 papers

For a general optimal control problem for dynamical systems with hybrid dynamics, we study the dependency of the optimal cost and of the value function on the initial conditions, parameters, and perturbations. We show that upper and lower…

Optimization and Control · Mathematics 2021-12-22 Berk Altın , Ricardo G. Sanfelice

We consider the problem of nonlinear stochastic optimal control. This problem is thought to be fundamentally intractable owing to Bellman's "curse of dimensionality". We present a result that shows that repeatedly solving an open-loop…

Systems and Control · Electrical Eng. & Systems 2024-10-11 Mohamed Naveed Gul Mohamed , Suman Chakravorty , Raman Goyal , Ran Wang

We consider a class of closed loop stochastic optimal control problems in finite time horizon, in which the cost is an expectation conditional on the event that the process has not exited a given bounded domain. An important difficulty is…

Optimization and Control · Mathematics 2019-12-19 Yves Achdou , Mathieu Laurière , Pierre-Louis Lions

We consider the problem of learning control policies that optimize a reward function while satisfying constraints due to considerations of safety, fairness, or other costs. We propose a new algorithm, Projection-Based Constrained Policy…

Machine Learning · Computer Science 2020-10-08 Tsung-Yen Yang , Justinian Rosca , Karthik Narasimhan , Peter J. Ramadge

This paper concerns discrete-time infinite-horizon stochastic control systems with Borel state and action spaces and universally measurable policies. We study optimization problems on strategic measures induced by the policies in these…

Optimization and Control · Mathematics 2023-12-22 Huizhen Yu

This work considers the stability of nonlinear stochastic receding horizon control when the optimal controller is only computed approximately. A number of general classes of controller approximation error are analysed including…

Optimization and Control · Mathematics 2018-12-03 Francesco Bertoli , Adrian N. Bishop

Given a discounted cost, we study deterministic discrete-time systems whose inputs are generated by policy iteration (PI). We provide novel near-optimality and stability properties, while allowing for non stabilizing initial policies. That…

Optimization and Control · Mathematics 2024-03-29 Jonathan de Brusse , Mathieu Granzotto , Romain Postoyan , Dragan Nešić

Many policies involve dynamics in their treatment assignments, where individuals receive sequential interventions over multiple stages. We study estimation of an optimal dynamic treatment regime that guides the optimal treatment assignment…

Econometrics · Economics 2024-09-04 Shosei Sakaguchi

A novel robust nonlinear model predictive control strategy is proposed for systems with nonlinear dynamics and convex state and control constraints. Using a sequential convex approximation approach and a difference of convex functions…

Optimization and Control · Mathematics 2025-01-28 Yana Lishkova , Mark Cannon

We study a multiscale stochastic optimal control problem subject to state constraints on the slow variable. To address this class of problems, we develop a rigorous theoretical framework based on singular perturbation analysis, tailored to…

Optimization and Control · Mathematics 2025-08-12 Anderson O. Calixto , Bernardo Freitas Paulo da Costa , Glauco Valle

We present a computer-checked generic implementation for solving finite-horizon sequential decision problems. This is a wide class of problems, including inter-temporal optimizations, knapsack, optimal bracketing, scheduling, etc. The…

Logic in Computer Science · Computer Science 2023-06-22 Nicola Botta , Patrik Jansson , Cezar Ionescu , David R. Christiansen , Edwin Brady

The Receding Horizon Control (RHC) strategy consists in replacing an infinite-horizon stabilization problem by a sequence of finite-horizon optimal control problems, which are numerically more tractable. The dynamic programming principle…

Optimization and Control · Mathematics 2019-06-06 Karl Kunisch , Laurent Pfeiffer

We study Markov decision processes with Polish state and action spaces. The action space is state dependent and is not necessarily compact. We first establish the existence of an optimal ergodic occupation measure using only a near-monotone…

Optimization and Control · Mathematics 2023-08-15 Ari Arapostathis , Vivek S. Borkar

This paper presents sufficient conditions for optimal control of systems with dynamics given by a linear operator, in order to obtain an explicit solution to the Bellman equation that can be calculated in a distributed fashion. Further, the…

Optimization and Control · Mathematics 2025-06-19 David Ohlin , Richard Pates , Murat Arcak

Optimal control of bilinear systems has been a well-studied subject in the area of mathematical control. However, techniques for solving emerging optimal control problems involving an ensemble of structurally identical bilinear systems are…

Optimization and Control · Mathematics 2016-01-14 Shuo Wang , Jr-Shin Li

This paper develops a policy learning method for tuning a pre-trained policy to adapt to additional tasks without altering the original task. A method named Adaptive Policy Gradient (APG) is proposed in this paper, which combines Bellman's…

Machine Learning · Computer Science 2025-09-29 Wenjian Hao , Zehui Lu , Zihao Liang , Tianyu Zhou , Shaoshuai Mou

This article is the starting point of a series of works whose aim is the study of deterministic control problems where the dynamic and the running cost can be completely different in two (or more) complementary domains of the space $\R^N$.…

Analysis of PDEs · Mathematics 2012-09-12 Guy Barles , Ariela Briani , Emmanuel Chasseigne

We consider the problem of sequential sampling from a finite number of independent statistical populations to maximize the expected infinite horizon average outcome per period, under a constraint that the expected average sampling cost does…

Machine Learning · Statistics 2012-01-20 Apostolos Burnetas , Odysseas Kanavetas

We provide a stability and performance analysis for nonlinear model predictive control (NMPC) schemes subject to input constraints. Given an exponential stabilizability and detectability condition w.r.t. the employed state cost, we provide…

Optimization and Control · Mathematics 2023-01-09 Johannes Köhler , Melanie N. Zeilinger , Lars Grüne

We present a novel particle filtering framework for continuous-time dynamical systems with continuous-time measurements. Our approach is based on the duality between estimation and optimal control, which allows reformulating the estimation…

Optimization and Control · Mathematics 2021-10-08 Qinsheng Zhang , Amirhossein Taghvaei , Yongxin Chen