English
Related papers

Related papers: Single Time-scale Actor-critic Method to Solve the…

200 papers

System stabilization via policy gradient (PG) methods has drawn increasing attention in both control and machine learning communities. In this paper, we study their convergence and sample complexity for stabilizing linear time-invariant…

Optimization and Control · Mathematics 2023-09-15 Feiran Zhao , Xingyun Fu , Keyou You

The present paper develops an optimal linear quadratic boundary controller for $2\times2$ linear hyperbolic partial differential equations (PDEs) with actuation on only one end of the domain. First-order necessary conditions for optimality…

Optimization and Control · Mathematics 2016-03-30 Agus Hasan , Lars Imsland , Ivan Ivanov , Snezhana Kostova , Boryana Bogdanova

We present and solve a Linear Quadratic Regulator (LQR) for the boundary control of the beam equation. We use the simple technique of completing the square to get an explicit solution. By decoupling the spatial frequencies we are able to…

Optimization and Control · Mathematics 2021-02-23 Arthur J. Krener

In cooperative stochastic games multiple agents work towards learning joint optimal actions in an unknown environment to achieve a common goal. In many real-world applications, however, constraints are often imposed on the actions that can…

Multiagent Systems · Computer Science 2020-07-14 Raghuram Bharadwaj Diddigi , Sai Koti Reddy Danda , Prabuchandran K. J. , Shalabh Bhatnagar

We propose a novel numerical method for high dimensional Hamilton--Jacobi--Bellman (HJB) type elliptic partial differential equations (PDEs). The HJB PDEs, reformulated as optimal control problems, are tackled by the actor-critic framework…

Optimization and Control · Mathematics 2022-01-07 Mo Zhou , Jiequn Han , Jianfeng Lu

Stochastic optimal control usually requires an explicit dynamical model with probability distributions, which are difficult to obtain in practice. In this work, we consider the linear quadratic regulator (LQR) problem of unknown linear…

Optimization and Control · Mathematics 2023-01-18 Feiran Zhao , Keyou You

On error of value function inevitably causes an overestimation phenomenon and has a negative impact on the convergence of the algorithms. To mitigate the negative effects of the approximation error, we propose Error Controlled Actor-critic…

Machine Learning · Computer Science 2021-09-08 Xingen Gao , Fei Chao , Changle Zhou , Zhen Ge , Chih-Min Lin , Longzhi Yang , Xiang Chang , Changjing Shang

A linear-quadratic (LQ, for short) optimal control problem is considered for mean-field stochastic differential equations with constant coefficients in an infinite horizon. The stabilizability of the control system is studied followed by…

Optimization and Control · Mathematics 2012-08-28 Jianhui Huang , Xun Li , Jiongmin Yong

This paper studies an infinite horizon optimal control problem for discrete-time linear system and quadratic criteria, both with random parameters which are independent and identically distributed with respect to time. In this general…

Optimization and Control · Mathematics 2024-03-04 Deyue Li

Actor-critic methods for decentralized multi-agent reinforcement learning (MARL) facilitate collaborative optimal decision making without centralized coordination, thus enabling a wide range of applications in practice. To date, however,…

Machine Learning · Computer Science 2025-08-14 Zhiyao Zhang , Myeung Suk Oh , FNU Hairi , Ziyue Luo , Alvaro Velasquez , Jia Liu

In this work, we consider policy-based methods for solving the reinforcement learning problem, and establish the sample complexity guarantees. A policy-based algorithm typically consists of an actor and a critic. We consider using various…

Machine Learning · Computer Science 2023-01-16 Zaiwei Chen , Siva Theja Maguluri

In this work, we propose a feedback control based temporal discretization for linear quadratic optimal control problems (LQ problems) governed by controlled mean-field stochastic differential equations. We firstly decompose the original…

Optimization and Control · Mathematics 2023-02-08 Yanqing Wang

This paper proposes a novel lifting method which converts the standard discrete-time linear periodic system to an augmented linear time-invariant system. The linear quadratic optimal control is then based on the solution of the…

Optimization and Control · Mathematics 2018-06-21 Yaguang Yang

Policy gradient algorithms are widely used in reinforcement learning and belong to the class of approximate dynamic programming methods. This paper studies two key policy gradient algorithms, the Natural Policy Gradient and the Gauss-Newton…

Systems and Control · Electrical Eng. & Systems 2026-05-11 Bowen Song , Sebastien Gros , Andrea Iannelli

Despite its nonconvexity, policy optimization for the Linear Quadratic Regulator (LQR) admits a favorable structural property known as gradient dominance, which facilitates linear convergence of policy gradient methods to the globally…

Optimization and Control · Mathematics 2026-02-27 Yuto Watanabe , Yang Zheng

We propose controller synthesis for state regulation problems in which a human operator shares control with an autonomy system, running in parallel. The autonomy system continuously improves over human action, with minimal intervention, and…

Systems and Control · Computer Science 2019-09-23 Murad Abu-Khalaf , Sertac Karaman , Daniela Rus

This paper presents an algorithm to solve the infinite horizon constrained linear quadratic regulator (CLQR) problem using operator splitting methods. First, the CLQR problem is reformulated as a (finite-time) model predictive control (MPC)…

Optimization and Control · Mathematics 2016-09-20 L. Ferranti , G. Stathopoulos , C. N. Jones , T. Keviczky

This paper focuses on adaptive control of the discrete-time linear quadratic regulator (adaptive LQR). Recent literature has made significant contributions in proving non-asymptotic convergence rates, but existing approaches have a few…

Systems and Control · Electrical Eng. & Systems 2026-04-27 Peter A. Fisher , Anuradha M. Annaswamy

Understanding the optimization landscape of linear quadratic regulation (LQR) problems is fundamental to the design of efficient reinforcement learning solutions. Recent work has made significant progress in characterizing the landscape of…

Systems and Control · Electrical Eng. & Systems 2026-04-14 Jingliang Duan , Jie Li , Yinsong Ma , Liye Tang , Guofa Li , Liping Zhang , Shengbo Eben Li , Lin Zhao

In this paper, we present a probability one convergence proof, under suitable conditions, of a certain class of actor-critic algorithms for finding approximate solutions to entropy-regularized MDPs using the machinery of stochastic…

Machine Learning · Computer Science 2019-10-23 Wesley Suttle , Zhuoran Yang , Kaiqing Zhang , Ji Liu
‹ Prev 1 3 4 5 6 7 10 Next ›