English
Related papers

Related papers: Receding-Horizon Policy Gradient for Polytopic Con…

200 papers

We develop two compression based stochastic gradient algorithms to solve a class of non-smooth strongly convex-strongly concave saddle-point problems in a decentralized setting (without a central server). Our first algorithm is a…

Machine Learning · Computer Science 2023-04-17 Chhavi Sharma , Vishnu Narayanan , P. Balamurugan

We study Concave Constrained Markov Decision Processes (Concave CMDPs) where both the objective and constraints are defined as concave functions of the state-action occupancy measure. We propose the Variance-Reduced Primal-Dual Policy…

Machine Learning · Computer Science 2024-05-28 Donghao Ying , Mengzi Amy Guo , Hyunin Lee , Yuhao Ding , Javad Lavaei , Zuo-Jun Max Shen

Adaptive Horizon Model Predictive Control (AHMPC) is a scheme for varying as needed the horizon length of Model Predictive Control (MPC). Its goal is to achieve stabilization with horizons as small as possible so that MPC can be used on…

Optimization and Control · Mathematics 2016-03-01 Arthur J. Krener

We begin an investigation of hybridizable discontinuous Galerkin (HDG) methods for approximating the solution of Dirichlet boundary control problems governed by elliptic PDEs. These problems can involve atypical variational formulations,…

Numerical Analysis · Mathematics 2017-12-11 Weiwei Hu , Jiguang Shen , John R. Singler , Yangwen Zhang , Xiaobo Zheng

We prove convergence of the proximal policy gradient method for a class of constrained stochastic control problems with control in both the drift and diffusion of the state process. The problem requires either the running or terminal cost…

Optimization and Control · Mathematics 2025-05-27 Ashley Davey , Harry Zheng

We study regret minimization for infinite-horizon average-reward Markov Decision Processes (MDPs) under cost constraints. We start by designing a policy optimization algorithm with carefully designed action-value estimator and bonus term,…

Machine Learning · Computer Science 2022-02-02 Liyu Chen , Rahul Jain , Haipeng Luo

We propose a scalable, policy-centric framework for continuous-time multi-asset portfolio-consumption optimization under inequality constraints. Our method integrates neural policies with Pontryagin's Maximum Principle (PMP) and enforces…

Portfolio Management · Quantitative Finance 2025-11-07 Jeonggyu Huh , Jaegi Jeon , Hyeng Keun Koo , Byung Hwa Lim

It is desirable but challenging to fulfill system constraints and reach optimal performance in consensus protocol design for practical multi-agent systems (MASs). This paper investigates the optimal consensus problem for general linear MASs…

Optimization and Control · Mathematics 2016-02-01 Huiping Li , Weisheng Yan , Yang Shi , Fuqiang Liu

Recently there has been an increasing interest in primal-dual methods for model predictive control (MPC), which require minimizing the (augmented) Lagrangian at each iteration. We propose a novel first order primal-dual method, termed…

Optimization and Control · Mathematics 2020-12-21 Yue Yu , Purnanand Elango , Behçet Açikmeşe

We develop a high-order hybridized discontinuous Galerkin (HDG) method for a linear degenerate elliptic equation arising from a two-phase mixture of mantle convection or glacier dynamics. We show that the proposed HDG method is well-posed…

Computational Engineering, Finance, and Science · Computer Science 2019-05-01 Shinhoo Kang , Tan Bui-Thanh , Todd Arbogast

We propose a robust model predictive control (MPC) method for discrete-time linear time-invariant systems with norm-bounded additive disturbances and model uncertainty. In our method, at each time step we solve a finite time robust optimal…

Systems and Control · Electrical Eng. & Systems 2021-11-11 Shaoru Chen , Nikolai Matni , Manfred Morari , Victor M. Preciado

This article proposes a distributed secondary control scheme that drives a dc microgrid to an equilibrium point where the generators share optimal currents, and their voltages have a weighted average of nominal value. The scheme does not…

Optimization and Control · Mathematics 2023-01-23 Babak Abdolmaleki , Gilbert Bergna-Diaz

This paper addresses the optimal control problem of finite-horizon discrete-time nonlinear systems under state and control constraints. A novel numerical algorithm based on optimal control theory is proposed to achieve superior…

Optimization and Control · Mathematics 2025-03-21 Chuanzhi Lv , Hongdan Li , Huanshui Zhang

We consider a finite-horizon linear-quadratic optimal control problem where only a limited number of control messages are allowed for sending from the controller to the actuator. To restrict the number of control actions computed and…

Systems and Control · Computer Science 2017-01-19 Burak Demirel , Euhanna Ghadimi , Daniel E. Quevedo , Mikael Johansson

This work considers the problem of computing the canonical polyadic decomposition (CPD) of large tensors. Prior works mostly leverage data sparsity to handle this problem, which is not suitable for handling dense tensors that often arise in…

Signal Processing · Electrical Eng. & Systems 2020-03-26 Xiao Fu , Shahana Ibrahim , Hoi-To Wai , Cheng Gao , Kejun Huang

Direct policy gradient methods for reinforcement learning and continuous control problems are a popular approach for a variety of reasons: 1) they are easy to implement without explicit knowledge of the underlying model 2) they are an…

Machine Learning · Computer Science 2019-03-26 Maryam Fazel , Rong Ge , Sham M. Kakade , Mehran Mesbahi

We study the problem of persistent monitoring of a finite number of inter-connected geographical nodes by a group of heterogeneous mobile agents. We assign to each geographical node a concave and increasing reward function that resets to…

Multiagent Systems · Computer Science 2020-10-22 Navid Rezazadeh , Solmaz S. Kia

Reinforcement Learning (RL) has emerged as a powerful framework for sequential decision-making in dynamic environments, particularly when system parameters are unknown. This paper investigates RL-based control for entropy-regularized…

Systems and Control · Electrical Eng. & Systems 2025-12-02 Gabriel Diaz , Lucky Li , Wenhao Zhang

We consider a distributed optimal control problem governed by an elliptic convection diffusion PDE, and propose a hybridizable discontinuous Galerkin (HDG) method to approximate the solution. We use polynomials of degree $k+1$ and $k \ge 0$…

Numerical Analysis · Mathematics 2018-11-27 Weiwei Hu , Jiguang Shen , John R. Singler , Yangwen Zhang , Xiaobo Zheng

To facilitate efficient learning, policy gradient approaches to deep reinforcement learning (RL) are typically paired with variance reduction measures and strategies for making large but safe policy changes based on a batch of experiences.…

Machine Learning · Computer Science 2023-11-13 Jared Markowitz , Edward W. Staley