English
Related papers

Related papers: Convergence Guarantees of Policy Optimization Meth…

200 papers

This paper is concerned with a stochastic linear-quadratic optimal control problem of Markovian regime switching system with model uncertainty and partial information, where the information available to the control is based on a…

Optimization and Control · Mathematics 2026-01-09 Na Xiang , Jingtao Shi

In this paper, we study a linear-quadratic optimal control problem for mean-field stochastic differential equations driven by a Poisson random martingale measure and a multidimensional Brownian motion. Firstly, the existence and uniqueness…

Optimization and Control · Mathematics 2016-10-12 Maoning Tang , Qingxin Meng

The convergence of policy gradient algorithms in reinforcement learning hinges on the optimization landscape of the underlying optimal control problem. Theoretical insights into these algorithms can often be acquired from analyzing those of…

Machine Learning · Computer Science 2023-11-01 Jingliang Duan , Wenhan Cao , Yang Zheng , Lin Zhao

This work presents a novel version of recently developed Gauss-Newton method for solving systems of nonlinear equations, based on upper bound of solution residual and quadratic regularization ideas. We obtained for such method global…

Optimization and Control · Mathematics 2021-05-04 Nikita Yudin , Alexander Gasnikov

We address the problem of finding an optimal policy in a Markov decision process under a restricted policy class defined by the convex hull of a set of base policies. This problem is of great interest in applications in which a number of…

Machine Learning · Computer Science 2018-02-28 Ershad Banijamali , Yasin Abbasi-Yadkori , Mohammad Ghavamzadeh , Nikos Vlassis

This paper investigates methods for estimating the optimal stochastic control policy for a Markov Decision Process with unknown transition dynamics and an unknown reward function. This form of model-free reinforcement learning comprises…

Machine Learning · Computer Science 2019-12-06 Brandon Trabucco , Albert Qu , Simon Li , Ganeshkumar Ashokavardhanan

We study the synthesis of mode switching protocols for a class of discrete-time switched linear systems in which the mode jumps are governed by Markov decision processes (MDPs). We call such systems MDP-JLS for brevity. Each state of the…

Systems and Control · Electrical Eng. & Systems 2020-01-06 Bo Wu , Murat Cubuktepe , Franck Djeumou , Zhe Xu , Ufuk Topcu

Reinforcement learning is a powerful tool to learn the optimal policy of possibly multiple agents by interacting with the environment. As the number of agents grow to be very large, the system can be approximated by a mean-field problem.…

Optimization and Control · Mathematics 2020-08-18 Weichen Wang , Jiequn Han , Zhuoran Yang , Zhaoran Wang

In this paper, we consider reinforcement learning of Markov Decision Processes (MDP) with peak constraints, where an agent chooses a policy to optimize an objective and at the same time satisfy additional constraints. The agent has to take…

Optimization and Control · Mathematics 2019-12-09 Ather Gattami

In this paper, we study the stabilization of two interdependent Markov jump linear systems (MJLSs) with partial information, where the interdependency arises as the transition of the mode of one system depends on the states of the other…

Systems and Control · Electrical Eng. & Systems 2020-05-26 Guanze Peng , Juntao Chen , Quanyan Zhu

Motivated by recent advances of reinforcement learning and direct data-driven control, we propose policy gradient adaptive control (PGAC) for the linear quadratic regulator (LQR), which uses online closed-loop data to improve the control…

Optimization and Control · Mathematics 2025-06-16 Feiran Zhao , Alessandro Chiuso , Florian Dörfler

We study the problem of computing deterministic optimal policies for constrained Markov decision processes (MDPs) with continuous state and action spaces, which are widely encountered in constrained dynamical systems. Designing…

Artificial Intelligence · Computer Science 2025-04-07 Sergio Rozada , Dongsheng Ding , Antonio G. Marques , Alejandro Ribeiro

This paper first describes a class of uncertain stochastic control systems with Markovian switching, and derives an It\^o-Liu formula for Markov-modulated processes. And we characterize an optimal control law, which satisfies the…

Optimization and Control · Mathematics 2014-01-14 Weiyin Fei

This paper addresses the stabilization problem of stochastic jump systems (SJSs) closed by a generally sampled controller. Because of the controller's switching and state both sampled, it is challenging to study its stabilization. A new…

Optimization and Control · Mathematics 2024-07-09 Guoliang Wang

Jump Markov linear systems (JMLS) are a useful class which can be used to model processes which exhibit random changes in behavior during operation. This paper presents a numerically stable method for learning the parameters of jump Markov…

Applications · Statistics 2020-04-21 Mark P. Balenzuela , Adrian G. Wills , Christopher Renton , Brett Ninness

We present a method to find an optimal policy with respect to a reward function for a discounted Markov decision process under general linear temporal logic (LTL) specifications. Previous work has either focused on maximizing a cumulative…

Systems and Control · Electrical Eng. & Systems 2021-03-24 Krishna C. Kalagarla , Rahul Jain , Pierluigi Nuzzo

We investigate the problem of optimal control synthesis for Markov Decision Processes (MDPs), addressing both qualitative and quantitative objectives. Specifically, we require the system to satisfy a qualitative task specified by a Linear…

Systems and Control · Electrical Eng. & Systems 2025-09-19 Yu Chen , Xuanyuan Yin , Shaoyuan Li , Xiang Yin

Machine learning techniques have demonstrated their effectiveness in achieving autonomy and optimality for nonlinear and high-dimensional dynamical systems. However, traditional black-box machine learning methods often lack formal stability…

Systems and Control · Electrical Eng. & Systems 2025-01-03 Kun Wang , Roberto Armellin , Adam Evans , Harry Holt , Zheng Chen

Infinite-time nonlinear optimal regulation control is widely utilized in aerospace engineering as a systematic method for synthesizing stable controllers. However, conventional methods often rely on linearization hypothesis, while recent…

Systems and Control · Electrical Eng. & Systems 2025-06-13 Han Wang , Di Wu , Lin Cheng , Shengping Gong , Xu Huang

Optimal feedback control with implicit Hamiltonians poses a fundamental challenge for learning-based value function methods due to the absence of closed-form optimal control laws. Recent work~\cite{gelphman2025end} introduced an implicit…

Optimization and Control · Mathematics 2026-04-28 Eric Gelphman , Deepanshu Verma , Nicole Tianjiao Yang , Stanley Osher , Samy Wu Fung