English
Related papers

Related papers: Split optimal policy iteration for LQR problems

200 papers

This paper proposes and analyzes two new policy learning methods: regularized policy gradient (RPG) and iterative policy optimization (IPO), for a class of discounted linear-quadratic control (LQC) problems over an infinite time horizon…

Optimization and Control · Mathematics 2025-10-08 Xin Guo , Xinyu Li , Renyuan Xu

We establish a Pontryagin maximum principle for discrete time optimal control problems under the following three types of constraints: a) constraints on the states pointwise in time, b) constraints on the control actions pointwise in time,…

Optimization and Control · Mathematics 2019-05-27 Pradyumna Paruchuri , Debasish Chatterjee

In this paper, we consider the inverse optimal control problem for the discrete-time linear quadratic regulator, over finite-time horizons. Given observations of the optimal trajectories, and optimal control inputs, to a linear…

Optimization and Control · Mathematics 2018-10-31 Han Zhang , Jack Umenberger , Xiaoming Hu

With the outstanding performance of policy gradient (PG) method in the reinforcement learning field, the convergence theory of it has aroused more and more interest recently. Meanwhile, the significant importance and abundant theoretical…

Optimization and Control · Mathematics 2024-04-19 Xinpei Zhang , Guangyan Jia

The problem of robust distributed control arises in several large-scale systems, such as transportation networks and power grid systems. In many practical scenarios controllers might not have enough information to make globally optimal…

Systems and Control · Computer Science 2019-09-26 Luca Furieri , Maryam Kamgarpour

In this paper, we study the local linear convergence properties of a versatile class of Primal-Dual splitting methods for minimizing composite non-smooth convex optimization problems. Under the assumption that the non-smooth components of…

Optimization and Control · Mathematics 2018-01-10 Jingwei Liang , Jalal Fadili , Gabriel Peyré

The main goal of this paper is to apply the so-called policy iteration algorithm (PIA) for the long run average continuous control problem of piecewise deterministic Markov processes (PDMP's) taking values in a general Borel space and with…

Probability · Mathematics 2009-02-17 O. L. V. Costa , F. Dufour

We consider linear iterative schemes for the time-discrete equations stemming from a class of nonlinear, doubly-degenerate parabolic equations. More precisely, the diffusion is nonlinear and may vanish or become multivalued for certain…

Numerical Analysis · Mathematics 2025-08-12 Ayesha Javed , Koondanibha Mitra , Iuliu Sorin Pop

We study stochastic optimal control problems for (possibly degenerate) McKean-Vlasov controlled diffusions and obtain discrete-time as well as finite interacting particle approximations. (i) Under mild assumptions, we first prove the…

Optimization and Control · Mathematics 2025-10-27 Somnath Pradhan , Serdar Yuksel

In safety-critical applications, reinforcement learning (RL) needs to consider safety constraints. However, theoretical understandings of constrained RL for continuous control are largely absent. As a case study, this paper presents a…

Optimization and Control · Mathematics 2024-06-07 Feiran Zhao , Keyou You

Combinatorial optimization is considered a promising class of problems in which quantum computers can show significant advantages. However, problems of practical relevance typically have more variables than current or foreseeable quantum…

Quantum Physics · Physics 2025-12-23 Mathias Schmid , Naeimeh Mohseni , Michael J. Hartmann

This paper studies policy transfer, one of the well-known transfer learning techniques adopted in large language models, for continuous-time reinforcement learning problems. In the case of continuous-time linear-quadratic systems with…

Machine Learning · Computer Science 2026-03-04 Xin Guo , Zijiu Lyu

This study presents the design, discretization and implementation of the continuous-time linear-quadratic model predictive control (CT-LMPC). The control model of the CT-LMPC is parameterized as transfer functions with time delays, and they…

Optimization and Control · Mathematics 2025-03-18 Zhanhao Zhang , Anders Hilmar Damm Christensen , Steen Hørsholt , John Bagterp Jørgensen

Inspired by the successes of stochastic algorithms in the training of deep neural networks and the simulation of interacting particle systems, we propose and analyze a framework for randomized time-splitting in linear-quadratic optimal…

Optimization and Control · Mathematics 2022-06-02 Daniel Veldman , Enrique Zuazua

This note introduces a new analytic approach to the solution of a very general class of finite-horizon optimal control problems formulated for discrete-time systems. This approach provides a parametric expression for the optimal control…

Optimization and Control · Mathematics 2012-09-03 Augusto Ferrante , Lorenzo Ntogramatzidis

This paper studies distributed algorithms for the extended monotropic optimization problem, which is a general convex optimization problem with a certain separable structure. The considered objective function is the sum of local convex…

Optimization and Control · Mathematics 2016-08-04 Xianlin Zeng , Peng Yi , Yiguang Hong , Lihua Xie

Dual decomposition is widely utilized in distributed optimization of multi-agent systems. In practice, the dual decomposition algorithm is desired to admit an asynchronous implementation due to imperfect communication, such as time delay…

Optimization and Control · Mathematics 2021-03-05 Yifan Su , Zhaojian Wang , Ming Cao , Mengshuo Jia , Feng Liu

Continuing the first part of the paper, we consider scalar decentralized average-cost infinite-horizon LQG problems with two controllers. This paper focuses on the slow dynamics case when the eigenvalue of the system is small and prove that…

Optimization and Control · Mathematics 2013-08-26 Se Yong Park , Anant Sahai

The convergence of policy gradient algorithms in reinforcement learning hinges on the optimization landscape of the underlying optimal control problem. Theoretical insights into these algorithms can often be acquired from analyzing those of…

Machine Learning · Computer Science 2023-11-01 Jingliang Duan , Wenhan Cao , Yang Zheng , Lin Zhao

This paper employs a policy iteration reinforcement learning (RL) method to study continuous-time linear-quadratic mean-field control problems in infinite horizon. The drift and diffusion terms in the dynamics involve the states, the…

Optimization and Control · Mathematics 2024-11-05 Na Li , Xun Li , Zuo Quan Xu
‹ Prev 1 8 9 10 Next ›