English
Related papers

Related papers: Ergodic-Risk Constrained Policy Optimization: The …

200 papers

The risk-neutral LQR controller is optimal for stochastic linear dynamical systems. However, the classical optimal controller performs inefficiently in the presence of low-probability yet statistically significant (risky) events. The…

Systems and Control · Electrical Eng. & Systems 2023-07-17 Masoud Roudneshin , Saba Sanami , Amir G. Aghdam

The linear quadratic regulator (LQR) problem has reemerged as an important theoretical benchmark for reinforcement learning-based control of complex dynamical systems with continuous state and action spaces. In contrast with nearly all…

Machine Learning · Computer Science 2020-05-04 Benjamin Gravell , Peyman Mohajerin Esfahani , Tyler Summers

This paper studies the robustness of reinforcement learning algorithms to errors in the learning process. Specifically, we revisit the benchmark problem of discrete-time linear quadratic regulation (LQR) and study the long-standing open…

Optimization and Control · Mathematics 2021-03-16 Bo Pang , Zhong-Ping Jiang

An optimal ergodic control problem (EC problem, for short) is investigated for a linear stochastic differential equation with quadratic cost functional. Constant nonhomogeneous terms, not all zero, appear in the state equation, which lead…

Optimization and Control · Mathematics 2020-04-24 Hongwei Mei , Qingmeng Wei , Jiongmin Yong

We formulate and solve a discrete-time linear-quadratic regulation (LQR) problem in a finite horizon that penalizes temporal variability and stochastic variability of the state trajectory. Our approach enables the user to strike a balance…

Optimization and Control · Mathematics 2026-03-26 Chuanning Wei , Kin Fung Li , Dionysis Kalogerias , Margaret P. Chapman

End-to-end engineering design pipelines, in which designs are evaluated using concurrently defined optimal controllers, are becoming increasingly common in practice. To discover designs that perform well even under the misspecification of…

Systems and Control · Electrical Eng. & Systems 2025-10-10 Yash Patel , Sahana Rayan , Ambuj Tewari

In ergodic singular stochastic control problems, a decision-maker can instantaneously adjust the evolution of a state variable using a control of bounded variation, with the goal of minimizing a long-term average cost functional. The cost…

Optimization and Control · Mathematics 2025-10-14 Alessandro Calvia , Federico Cannerozzi , Giorgio Ferrari

We investigate the ergodic problem of growth-rate maximization under a class of risk constraints in the context of incomplete, It\^{o}-process models of financial markets with random ergodic coefficients. Including {\em value-at-risk}…

Portfolio Management · Quantitative Finance 2008-12-02 Traian A. Pirvu , Gordan Zitkovic

This paper focuses on the linear quadratic control (LQC) design of systems corrupted by both stochastic noise and bounded noise simultaneously. When only of these noises are considered, the LQC strategy leads to stochastic or robust…

Optimization and Control · Mathematics 2025-12-15 Xuehui Ma , Shiliang Zhang , Xiaohui Zhang , Jing Xin , Hector Garcia de Marina

This paper presents a convex optimization-based solution to the design of state-feedback controllers for solving the linear quadratic regulator (LQR) problem of uncertain discrete-time systems with multiplicative noise. To synthesize a…

Systems and Control · Electrical Eng. & Systems 2022-05-17 Majid Mazouchi , Farzaneh Tatari , Hamidreza Modares

This paper proposes and analyzes two new policy learning methods: regularized policy gradient (RPG) and iterative policy optimization (IPO), for a class of discounted linear-quadratic control (LQC) problems over an infinite time horizon…

Optimization and Control · Mathematics 2025-10-08 Xin Guo , Xinyu Li , Renyuan Xu

Having a perfect model to compute the optimal policy is often infeasible in reinforcement learning. It is important in high-stakes domains to quantify and manage risk induced by model uncertainties. Entropic risk measure is an exponential…

Machine Learning · Computer Science 2020-06-23 Reazul Hasan Russel , Bahram Behzadian , Marek Petrik

Despite the empirical success of the actor-critic algorithm, its theoretical understanding lags behind. In a broader context, actor-critic can be viewed as an online alternating update algorithm for bilevel optimization, whose convergence…

Machine Learning · Computer Science 2019-07-16 Zhuoran Yang , Yongxin Chen , Mingyi Hong , Zhaoran Wang

This paper addresses a risk-constrained decentralized stochastic linear-quadratic optimal control problem with one remote controller and one local controller, where the risk constraint is posed on the cumulative state weighted variance in…

Optimization and Control · Mathematics 2023-07-19 Jia Hui , Yuan-Hua Ni

This paper proposes a novel robust reinforcement learning framework for discrete-time linear systems with model mismatch that may arise from the sim-to-real gap. A key strategy is to invoke advanced techniques from control theory. Using the…

Systems and Control · Electrical Eng. & Systems 2023-12-07 Leilei Cui , Tamer Başar , Zhong-Ping Jiang

Ergodicity describes an equivalence between the expectation value and the time average of observables. Applied to human behaviour, ergodic theories of decision-making reveal how individuals should tolerate risk in different environments. To…

The Linear Quadratic Gaussian (LQG) regulator is a cornerstone of optimal control theory, yet its performance can degrade significantly when the noise distributions deviate from the assumed Gaussian model. To address this limitation, this…

Systems and Control · Electrical Eng. & Systems 2026-03-27 Riccardo Cescon , Andrea Martin , Giancarlo Ferrari-Trecate

The Linear Quadratic Gaussian (LQG) controller is known to be inherently fragile to model misspecifications common in real-world situations. We consider discrete-time partially observable stochastic linear systems and provide a…

Optimization and Control · Mathematics 2025-07-31 Marta Fochesato , Lucia Falconi , Mattia Zorzi , Augusto Ferrante , John Lygeros

Iterative trajectory optimization techniques for non-linear dynamical systems are among the most powerful and sample-efficient methods of model-based reinforcement learning and approximate optimal control. By leveraging time-variant local…

Systems and Control · Electrical Eng. & Systems 2019-08-01 Onur Celik , Hany Abdulsamad , Jan Peters

We study the infinite-horizon average (ergodic) risk sensitive control problem for diffusion processes under a general structural hypothesis: there is a partition of state space into two subsets, where the controlled diffusion process…

Optimization and Control · Mathematics 2025-12-01 Sumith Reddy Anugu , Guodong Pang