English
Related papers

Related papers: Solving the Model Unavailable MARE using Q-Learnin…

200 papers

The development of machine learning algorithms has been gathering relevance to address the increasing modelling complexity of manufacturing decision-making problems. Reinforcement learning is a methodology with great potential due to the…

Machine Learning · Computer Science 2023-04-18 Miguel Neves , Miguel Vieira , Pedro Neto

In this paper, two Q-learning (QL) methods are proposed and their convergence theories are established for addressing the model-free optimal control problem of general nonlinear continuous-time systems. By introducing the Q-function for…

Systems and Control · Computer Science 2014-10-14 Biao Luo , Derong Liu , Tingwen Huang

The stochastic $H_{\infty}$ control is studied for a linear stochastic It\^o system with an unknown system model. The linear stochastic $H_{\infty}$ control issue is known to be transformable into the problem of solving a so-called…

Optimization and Control · Mathematics 2024-03-08 Jing Guo Jing Guo , Xiushan Jiang , Weihai Zhang

We revisit and extend the Riccati theory, unifying continuous-time linear-quadratic optimal permanent and sampled-data control problems, in finite and infinite time horizons. In a nutshell, we prove that:-- when the time horizon T tends to…

Optimization and Control · Mathematics 2020-02-12 Loïc Bourdin , Emmanuel Trélat

Linear-Quadratic (LQ) problems that arise in systems and controls include the classical optimal control problems of the Linear Quadratic Regulator (LQR) in both its deterministic and stochastic forms, as well as $H^\infty$-analysis (the…

Systems and Control · Electrical Eng. & Systems 2024-01-04 Bassam Bamieh

The asymptotic iteration method (AIM) is an iterative technique used to find exact and approximate solutions to second-order linear differential equations. In this work, we employed AIM to solve systems of two first-order linear…

Mathematical Physics · Physics 2009-01-15 Katherine M. Robertson , Nasser Saad

In standard linear quadratic (LQ) control, the first step in investigating infinite-horizon optimal control is to derive the stabilization condition with the optimal LQ controller. This paper focuses on the stabilization of an Ito…

Optimization and Control · Mathematics 2019-08-22 Hongdan Li , Qingyuan Qi , Huanshui Zhang

The QLBS model is a discrete-time option hedging and pricing model that is based on Dynamic Programming (DP) and Reinforcement Learning (RL). It combines the famous Q-Learning method for RL with the Black-Scholes (-Merton) model's idea of…

Computational Finance · Quantitative Finance 2018-01-19 Igor Halperin

A discrete-time stochastic LQ problem with multiplicative noises and state transmission delay is studied in this paper, which does not require any definiteness constraint on the cost weighting matrices. From some abstract representations of…

Optimization and Control · Mathematics 2017-05-30 Yuan-Hua Ni , Cedric Ka-Fai Yiu , Huanshui Zhang , Ji-Feng Zhang

This paper formulates a stochastic optimal control problem for linear networked control systems featuring stochastic packet disordering with a unique stabilizing solution certified. The problem is solved by proposing reinforcement learning…

Systems and Control · Electrical Eng. & Systems 2023-12-13 Wenqian Xue , Yi Jiang , Frank L. Lewis , Bosen Lian

The coupled Riccati equations are cosisted of multiple Riccati-like equations with solutions coupled with each other, which can be applied to depict the properties of more complex systems such as markovian systems or multi-agent systems.…

Signal Processing · Electrical Eng. & Systems 2023-07-14 Jiachen Qian , Peihu Duan , Zhisheng Duan , Ling shi

This paper proposes a reinforcement learning (RL) algorithm for infinite horizon $\rm {H_{2}/H_{\infty}}$ problem in a class of stochastic discrete-time systems, rather than using a set of coupled generalized algebraic Riccati equations…

Optimization and Control · Mathematics 2023-11-28 Xiushan Jiang , Li Wang , Dongya Zhao , Ling Shi

We consider the algebraic Riccati equation for which the four coefficient matrices form an M-matrix K. When K is a nonsingular M-matrix or an irreducible singular M-matrix, the Riccati equation is known to have a minimal nonnegative…

Numerical Analysis · Mathematics 2013-01-01 Chun-Hua Guo

Methods from learning theory are used in the state space of linear dynamical and control systems in order to estimate the system matrices. An application to stabilization via algebraic Riccati equations is included. The approach is…

Dynamical Systems · Mathematics 2015-08-12 Fritz Colonius , Boumediene Hamzi

Dynamic treatment regimes (DTRs) formalize medical decision-making as a sequence of rules for different stages, mapping patient-level information to recommended treatments. In practice, estimating an optimal DTR using observational data…

Methodology · Statistics 2024-12-11 Jian Sun , Bo Fu , Li Su

In this paper, the open-loop, closed-loop, and weak closed-loop solvability for discrete-time linear-quadratic (LQ) control problem is considered due to the fact that it is always open-loop optimal solvable if the LQ control problem is…

Optimization and Control · Mathematics 2025-02-18 Yue Sun , Xianping Wu , Xun Li

This paper introduces and analyzes an improved Q-learning algorithm for discrete-time linear time-invariant systems. The proposed method does not require any knowledge of the system dynamics, and it enjoys significant efficiency advantages…

Systems and Control · Electrical Eng. & Systems 2023-04-03 Victor G. Lopez , Mohammad Alsalti , Matthias A. Müller

We give a rank characterization of the solution set of algebraic Riccati inequality (ARI) for both controllable and uncontrollable systems. Assuming an existence of a solution of the corresponding algebraic Riccati equation (ARE), we…

Optimization and Control · Mathematics 2019-03-01 A. Sanand Amita Dilip , Harish K. Pillai

An iterative scheme for the Dynamical Systems Method (DSM) is given such that one does not have to solve the Cauchy problem occuring in the application of the DSM for solving ill-conditioned linear algebraic systems. The novelty of the…

Numerical Analysis · Mathematics 2008-03-25 N. S. Hoang , A. G. Ramm

Non-stationarity is a fundamental challenge in multi-agent reinforcement learning (MARL), where agents update their behaviour as they learn. Many theoretical advances in MARL avoid the challenge of non-stationarity by coordinating the…

Computer Science and Game Theory · Computer Science 2025-03-19 Bora Yongacoglu , Gürdal Arslan , Serdar Yüksel