中文
相关论文

相关论文: Solving the Model Unavailable MARE using Q-Learnin…

200 篇论文

This paper is concerned with a backward stochastic linear-quadratic (LQ, for short) optimal control problem with deterministic coefficients. The weighting matrices are allowed to be indefinite, and cross-product terms in the control and…

最优化与控制 · 数学 2021-04-13 Jingrui Sun , Zhen Wu , Jie Xiong

Reachability analysis is a fundamental problem for safety verification and falsification of Cyber-Physical Systems (CPS) whose dynamics follow physical laws usually represented as differential equations. In the last two decades, numerous…

符号计算 · 计算机科学 2018-04-11 Hoang-Dung Tran , Weiming Xiang , Nathaniel Hamilton , Taylor T. Johnson

Missing data is an important problem in machine learning practice. Starting from the premise that imputation methods should preserve the causal structure of the data, we develop a regularization scheme that encourages any baseline…

机器学习 · 计算机科学 2021-11-08 Trent Kyono , Yao Zhang , Alexis Bellot , Mihaela van der Schaar

A study of the linear quadratic (LQ) control problem on a finite time interval for a model equation in Hilbert spaces which comprehends the memory of the inputs was performed recently by the authors. The outcome included a closed-loop…

最优化与控制 · 数学 2025-03-19 Paolo Acquistapace , Francesca Bucci

We propose a novel offline reinforcement learning (offline RL) approach, introducing the Diffusion-model-guided Implicit Q-learning with Adaptive Revaluation (DIAR) framework. We address two key challenges in offline RL: out-of-distribution…

机器学习 · 计算机科学 2024-10-16 Jaehyun Park , Yunho Kim , Sejin Kim , Byung-Jun Lee , Sundong Kim

Regret analysis is challenging in Multi-Agent Reinforcement Learning (MARL) primarily due to the dynamical environments and the decentralized information among agents. We attempt to solve this challenge in the context of decentralized…

机器学习 · 计算机科学 2020-01-29 Seyed Mohammad Asghari , Yi Ouyang , Ashutosh Nayyar

In this paper, we focus on using optimization methods to solve matrix equations by transforming the problem of solving the Sylvester matrix equation or continuous algebraic Riccati equation into an optimization problem. Initially, we use a…

数值分析 · 数学 2024-04-10 Juan Zhang , Xiao Luo

This paper addresses the mean-square optimal control problem for \a class of discrete-time linear systems with a quasi-colored control-dependent multiplicative noise via output feedback. The noise under study is novel and shown to have…

系统与控制 · 电气工程与系统科学 2021-09-06 Junhui Li , Jieying Lu , Weizhou Su

Quantum annealing (QA) has great potential to solve combinatorial optimization problems efficiently. However, the effectiveness of QA algorithms is heavily based on the embedding of problem instances, represented as logical graphs, into the…

人工智能 · 计算机科学 2025-10-07 Hoang M. Ngo , Nguyen H K. Do , Minh N. Vu , Tre' R. Jeter , Tamer Kahveci , My T. Thai

Randomized iterative methods, such as the Kaczmarz method and its variants, have gained growing attention due to their simplicity and efficiency in solving large-scale linear systems. Meanwhile, absolute value equations (AVE) have attracted…

数值分析 · 数学 2025-05-13 Jiaxin Xie , Hou-Duo Qi , Deren Han

This paper studies data-driven approaches to the continuous-time linear quadratic regulator (LQR) problem based on two existing parameterizations, namely a closed-loop (CL) parameterization from behavioral system theory and an integral…

最优化与控制 · 数学 2026-05-01 Armin Gießler , Felix Thömmes , Sören Hohmann

The quasilinearization method (QLM) of solving nonlinear differential equations is applied to the quantum mechanics by casting the Schr\"{o}dinger equation in the nonlinear Riccati form. The method, whose mathematical basis in physics was…

计算物理 · 物理学 2007-05-23 R. Krivec , V. B. Mandelzweig

We identify a relationship between the solutions of a nonsymmetric algebraic T-Riccati equation (T-NARE) and the deflating subspaces of a palindromic matrix pencil, obtained by arranging the coefficients of the T-NARE. The interplay between…

数值分析 · 数学 2021-10-08 Peter Benner , Bruno Iannazzo , Beatrice Meini , Davide Palitta

This paper examines the nonconvex quadratically constrained quadratic programming (QCQP) problems using an iterative method. One of the existing approaches for solving nonconvex QCQP problems relaxes the rank one constraint on the unknown…

最优化与控制 · 数学 2016-09-12 Chuangchuang Sun , Ran Dai

Reinforcement learning has witnessed significant advancements, particularly with the emergence of model-based approaches. Among these, $Q$-learning has proven to be a powerful algorithm in model-free settings. However, the extension of…

机器学习 · 计算机科学 2026-03-31 Han-Dong Lim , HyeAnn Lee , Donghwan Lee

Value estimation is one key problem in Reinforcement Learning. Albeit many successes have been achieved by Deep Reinforcement Learning (DRL) in different fields, the underlying structure and learning dynamics of value function, especially…

机器学习 · 计算机科学 2021-11-22 Tong Sang , Hongyao Tang , Jianye Hao , Yan Zheng , Zhaopeng Meng

We propose a variational quantum eigensolver (VQE) for the simulation of strongly-correlated quantum matter based on a multi-scale entanglement renormalization ansatz (MERA) and gradient-based optimization. This MERA quantum eigensolver can…

量子物理 · 物理学 2023-09-04 Qiang Miao , Thomas Barthel

A consistent Riccati expansion (CRE) is proposed for solving nonlinear systems with the help of a Riccati equation. A system is defined to be CRE solvable if it has a CRE. Various integrable systems are CRE solvable. Furthermore, it is also…

可精确求解与可积系统 · 物理学 2013-09-02 S. Y. Lou

Deep reinforcement learning has been extensively studied in decision-making processes and has demonstrated superior performance over conventional approaches in various fields, including radar resource management (RRM). However, a notable…

机器学习 · 计算机科学 2025-06-27 Ziyang Lu , M. Cenk Gursoy , Chilukuri K. Mohan , Pramod K. Varshney

We investigate a class of zero-sum linear-quadratic stochastic differential games on a finite time horizon governed by multiscale state equations. The multiscale nature of the problem can be leveraged to reformulate the associated…

最优化与控制 · 数学 2020-11-19 Beniamin Goldys , James Yang , Zhou Zhou