中文
相关论文

相关论文: Solving the Model Unavailable MARE using Q-Learnin…

200 篇论文

Real economies can be modeled as a sequential imperfect-information game with many heterogeneous agents, such as consumers, firms, and governments. Dynamic general equilibrium (DGE) models are often used for macroeconomic analysis in this…

计算机科学与博弈论 · 计算机科学 2022-02-25 Michael Curry , Alexander Trott , Soham Phade , Yu Bai , Stephan Zheng

Discrete algebraic Riccati equations and their fixed points are well understood and arise in a variety of applications, however, the time-varying equations have not yet been fully explored in the literature. In this article we provide a…

动力系统 · 数学 2021-07-28 Pierre del Moral , Emma Horton

Model-free Reinforcement Learning (RL) algorithms such as Q-learning [Watkins, Dayan 92] have been widely used in practice and can achieve human level performance in applications such as video games [Mnih et al. 15]. Recently, equipped with…

机器学习 · 计算机科学 2019-05-03 Zhao Song , Wen Sun

A novel integrability condition for the Riccati equation, the simplest form of nonlinear ordinary differential equations, is obtained by using elementary quadrature method. Under this condition, the analytic general solution is presented,…

可精确求解与可积系统 · 物理学 2026-04-09 Zhao Ji-Xiang

Model-free approaches for reinforcement learning (RL) and continuous control find policies based only on past states and rewards, without fitting a model of the system dynamics. They are appealing as they are general purpose and easy to…

机器学习 · 计算机科学 2018-10-09 Yasin Abbasi-Yadkori , Nevena Lazic , Csaba Szepesvari

Finite-time linear-quadratic control of partial differential-algebraic equations (PDAEs) is considered. The discussion is restricted to those that are radial with index $0$; this corresponds to a nilpotency degree of 1. We establish the…

最优化与控制 · 数学 2024-04-08 Ala' Alalabi , Kirsten Morris

Recently, model-free reinforcement learning has attracted research attention due to its simplicity, memory and computation efficiency, and the flexibility to combine with function approximation. In this paper, we propose Exploration…

机器学习 · 计算机科学 2020-12-10 Mehdi Jafarnia-Jahromi , Chen-Yu Wei , Rahul Jain , Haipeng Luo

Efficient Riccati equation based techniques for the approximate solution of discrete time linear regulator problems are restricted in their application to problems with quadratic terminal payoffs. Where non-quadratic terminal payoffs are…

最优化与控制 · 数学 2017-11-13 Huan Zhang , Peter M. Dower

An indefinite stochastic Riccati Equation is a matrix-valued, highly nonlinear backward stochastic differential equation together with an algebraic, matrix positive definiteness constraint. We introduce a new approach to solve a class of…

概率论 · 数学 2012-03-20 Zhongmin Qian , Xun Yu Zhou

Reinforcement learning with offline data suffers from Q-value extrapolation errors. To address this issue, we first demonstrate that linear extrapolation of the Q-function beyond the data range is particularly problematic. To mitigate this,…

The Riccati differential equation is examined in light of its connection to second order linear time varying systems. In that light it becomes the clear generalization for the characteristic equation of linear time invariant systems, and is…

动力系统 · 数学 2026-04-24 Douglas R. Frey

This work aims to tackle a major challenge in offline Inverse Reinforcement Learning (IRL), namely the reward extrapolation error, where the learned reward function may fail to explain the task correctly and misguide the agent in unseen…

机器学习 · 计算机科学 2023-02-22 Sheng Yue , Guanbo Wang , Wei Shao , Zhaofeng Zhang , Sen Lin , Ju Ren , Junshan Zhang

In warehouses, specialized agents need to navigate, avoid obstacles and maximize the use of space in the warehouse environment. Due to the unpredictability of these environments, reinforcement learning approaches can be applied to complete…

In a recent work, we proposed a graph-based manifold learning scheme for the nonlinear Galerkin-reduction of quasi-static solid mechanical problems [1]. The resulting nonlinear approximation spaces can closely and flexibly represent…

计算工程、金融与科学 · 计算机科学 2025-09-01 Erik Faust , Lisa Scheunemann

This note concerns a class of matrix Riccati equations associated with stochastic linear-quadratic optimal control problems with indefinite state and control weighting costs. A novel sufficient condition of solvability of such equations is…

最优化与控制 · 数学 2013-12-30 Kai Du

We propose an SQP algorithm for mathematical programs with vanishing constraints which solves at each iteration a quadratic program with linear vanishing constraints. The algorithm is based on the newly developed concept of $\mathcal…

最优化与控制 · 数学 2016-11-28 Matúš Benko , Helmut Gfrerer

We present two matrix-free methods for approximately solving exact penalty subproblems that arise when solving large-scale optimization problems. The first approach is a novel iterative re-weighting algorithm (IRWA), which iteratively…

最优化与控制 · 数学 2017-01-02 James V. Burke , Frank E. Curtis , Hao Wang , Jiashan Wang

A new method, the Dynamical Systems Method (DSM), justified recently, is applied to solving ill-conditioned linear algebraic system (ICLAS). The DSM gives a new approach to solving a wide class of ill-posed problems. In this paper a new…

数值分析 · 数学 2009-12-04 Sapto W. Indratno , A. G. Ramm

Quality Diversity (QD) has emerged as a powerful alternative optimization paradigm that aims at generating large and diverse collections of solutions, notably with its flagship algorithm MAP-ELITES (ME) which evolves solutions through…

神经与进化计算 · 计算机科学 2023-06-16 Thomas Pierrot , Arthur Flajolet

While the Matrix Generalized Inverse Gaussian ($\mathcal{MGIG}$) distribution arises naturally in some settings as a distribution over symmetric positive semi-definite matrices, certain key properties of the distribution and effective ways…

机器学习 · 统计学 2016-08-23 Farideh Fazayeli , Arindam Banerjee