中文
相关论文

相关论文: Solving the Model Unavailable MARE using Q-Learnin…

200 篇论文

We propose a new algorithm for a broad class of periodic time-varying Stochastic Game-Theoretic Riccati Differential Equations arising in Zero-Sum Linear-Quadratic Stochastic Differential Games. The algorithm is constructed via dual-layer…

数值分析 · 数学 2025-11-06 Yiyuan Wang

Inspired by REINFORCE, we introduce a novel receding-horizon algorithm for the Linear Quadratic Regulator (LQR) problem with unknown dynamics. Unlike prior methods, our algorithm avoids reliance on two-point gradient estimates while…

最优化与控制 · 数学 2025-10-07 Amirreza Neshaei Moghaddam , Alex Olshevsky , Bahman Gharesifard

In this paper we discuss how to decompose the constrained generalized discrete-time algebraic Riccati equation arising in optimal control and optimal filtering problems into two parts corresponding to an additive decomposition X=X0+D of…

最优化与控制 · 数学 2014-09-24 Lorenzo Ntogramatzidis , Augusto Ferrante

Domain-specific intelligence demands specialized knowledge and sophisticated reasoning for problem-solving, posing significant challenges for large language models (LLMs) that struggle with knowledge hallucination and inadequate reasoning…

计算与语言 · 计算机科学 2025-05-20 Zhengren Wang , Jiayang Yu , Dongsheng Ma , Zhe Chen , Yu Wang , Zhiyu Li , Feiyu Xiong , Yanfeng Wang , Weinan E , Linpeng Tang , Wentao Zhang

Humans quickly solve tasks in novel systems with complex dynamics, without requiring much interaction. While deep reinforcement learning algorithms have achieved tremendous success in many complex tasks, these algorithms need a large number…

Functional relation for commuting quantum transfer matrices of quantum integrable models is identified with classical Hirota's bilinear difference equation. This equation is equivalent to the completely discretized classical 2D Toda lattice…

高能物理 - 理论 · 物理学 2019-08-15 I. Krichever , O. Lipan , P. Wiegmann , A. Zabrodin

This paper analyzes a special instance of nonsymmetric algebraic matrix Riccati equations arising from transport theory. Traditional approaches for finding the minimal nonnegative solution of the matrix Riccati equations are based on the…

数值分析 · 数学 2011-09-26 Chun-Yueh Chiang , Matthew M. Lin

This paper studies the robustness of reinforcement learning algorithms to errors in the learning process. Specifically, we revisit the benchmark problem of discrete-time linear quadratic regulation (LQR) and study the long-standing open…

最优化与控制 · 数学 2021-03-16 Bo Pang , Zhong-Ping Jiang

In this paper we consider a class of conjugate discrete-time Riccati equations, arising originally from the linear quadratic regulation problem for discrete-time antilinear systems. Under mild and reasonable assumptions, the existence of…

最优化与控制 · 数学 2022-12-06 Chun-Yueh Chiang , Hung-Yuan Fan

In this paper, the solvability of discrete-time stochastic linear-quadratic (LQ) optimal control problem in finite horizon is considered. Firstly, it shows that the closed-loop solvability for the LQ control problem is optimal if and only…

最优化与控制 · 数学 2025-02-25 Yue Sun , Xianping Wu , Xun Li

This paper presents a state and state-input constrained variant of the discrete-time iterative Linear Quadratic Regulator (iLQR) algorithm, with linear time-complexity in the number of time steps. The approach is based on a projection of…

机器人学 · 计算机科学 2018-05-25 Markus Giftthaler , Jonas Buchli

Algebraic Riccati equations (AREs) have been extensively applicable in linear optimal control problems and many efficient numerical methods were developed. The most attention of numerical solutions is the (almost) stabilizing solution in…

最优化与控制 · 数学 2021-11-18 Chun-Yueh Chiang , Hung-Yuan Fan

A novel algebra underlying integrable systems is shown to generate and unify a large class of quantum integrable models with given $R$-matrix, through reductions of an ancestor Lax operator and its different realizations. Along with known…

高能物理 - 理论 · 物理学 2009-10-31 Anjan Kundu

Unmanned Aerial Vehicles need an online path planning capability to move in high-risk missions in unknown and complex environments to complete them safely. However, many algorithms reported in the literature may not return reliable…

This paper presents a pioneering approach to solving the linear quadratic regulation (LQR) and linear quadratic tracking (LQT) problems with constrained inputs using a novel off-policy continuous-time Q-learning framework. The proposed…

系统与控制 · 电气工程与系统科学 2025-09-23 Duc Cuong Nguyen , Quang Huy Dao , Phuong Nam Dao

Multi-agent reinforcement learning (MARL) suffers from the non-stationarity problem, which is the ever-changing targets at every iteration when multiple agents update their policies at the same time. Starting from first principle, in this…

机器学习 · 计算机科学 2022-12-05 Chuming Li , Jie Liu , Yinmin Zhang , Yuhong Wei , Yazhe Niu , Yaodong Yang , Yu Liu , Wanli Ouyang

A class of differential Riccati equations (DREs) is considered whereby the evolution of any solution can be identified with the propagation of a value function of a corresponding optimal control problem arising in L2-gain analysis. By…

最优化与控制 · 数学 2017-11-13 Peter M. Dower , Huan Zhang

Q-learning is a popular reinforcement learning algorithm. This algorithm has however been studied and analysed mainly in the infinite horizon setting. There are several important applications which can be modeled in the framework of finite…

机器学习 · 计算机科学 2022-08-09 Vivek VP , Dr. Shalabh Bhatnagar

Differential algebraic Riccati equations are at the heart of many applications in control theory. They are time-depent, matrix-valued, and in particular nonlinear equations that require special methods for their solution. Low-rank methods…

数值分析 · 数学 2019-12-17 Tobias Breiten , Sergey Dolgov , Martin Stoll

Continuous-time algebraic Riccati equations can be found in many disciplines in different forms. In the case of small-scale dense coefficient matrices, stabilizing solutions can be computed to all possible formulations of the Riccati…

数值分析 · 数学 2024-09-18 Jens Saak , Steffen W. R. Werner