English
Related papers

Related papers: Solving the Model Unavailable MARE using Q-Learnin…

200 papers

Real economies can be modeled as a sequential imperfect-information game with many heterogeneous agents, such as consumers, firms, and governments. Dynamic general equilibrium (DGE) models are often used for macroeconomic analysis in this…

Computer Science and Game Theory · Computer Science 2022-02-25 Michael Curry , Alexander Trott , Soham Phade , Yu Bai , Stephan Zheng

Discrete algebraic Riccati equations and their fixed points are well understood and arise in a variety of applications, however, the time-varying equations have not yet been fully explored in the literature. In this article we provide a…

Dynamical Systems · Mathematics 2021-07-28 Pierre del Moral , Emma Horton

Model-free Reinforcement Learning (RL) algorithms such as Q-learning [Watkins, Dayan 92] have been widely used in practice and can achieve human level performance in applications such as video games [Mnih et al. 15]. Recently, equipped with…

Machine Learning · Computer Science 2019-05-03 Zhao Song , Wen Sun

A novel integrability condition for the Riccati equation, the simplest form of nonlinear ordinary differential equations, is obtained by using elementary quadrature method. Under this condition, the analytic general solution is presented,…

Exactly Solvable and Integrable Systems · Physics 2026-04-09 Zhao Ji-Xiang

Model-free approaches for reinforcement learning (RL) and continuous control find policies based only on past states and rewards, without fitting a model of the system dynamics. They are appealing as they are general purpose and easy to…

Machine Learning · Computer Science 2018-10-09 Yasin Abbasi-Yadkori , Nevena Lazic , Csaba Szepesvari

Finite-time linear-quadratic control of partial differential-algebraic equations (PDAEs) is considered. The discussion is restricted to those that are radial with index $0$; this corresponds to a nilpotency degree of 1. We establish the…

Optimization and Control · Mathematics 2024-04-08 Ala' Alalabi , Kirsten Morris

Recently, model-free reinforcement learning has attracted research attention due to its simplicity, memory and computation efficiency, and the flexibility to combine with function approximation. In this paper, we propose Exploration…

Machine Learning · Computer Science 2020-12-10 Mehdi Jafarnia-Jahromi , Chen-Yu Wei , Rahul Jain , Haipeng Luo

Efficient Riccati equation based techniques for the approximate solution of discrete time linear regulator problems are restricted in their application to problems with quadratic terminal payoffs. Where non-quadratic terminal payoffs are…

Optimization and Control · Mathematics 2017-11-13 Huan Zhang , Peter M. Dower

An indefinite stochastic Riccati Equation is a matrix-valued, highly nonlinear backward stochastic differential equation together with an algebraic, matrix positive definiteness constraint. We introduce a new approach to solve a class of…

Probability · Mathematics 2012-03-20 Zhongmin Qian , Xun Yu Zhou

Reinforcement learning with offline data suffers from Q-value extrapolation errors. To address this issue, we first demonstrate that linear extrapolation of the Q-function beyond the data range is particularly problematic. To mitigate this,…

Machine Learning · Computer Science 2025-08-20 Jeonghye Kim , Yongjae Shin , Whiyoung Jung , Sunghoon Hong , Deunsol Yoon , Youngchul Sung , Kanghoon Lee , Woohyung Lim

The Riccati differential equation is examined in light of its connection to second order linear time varying systems. In that light it becomes the clear generalization for the characteristic equation of linear time invariant systems, and is…

Dynamical Systems · Mathematics 2026-04-24 Douglas R. Frey

This work aims to tackle a major challenge in offline Inverse Reinforcement Learning (IRL), namely the reward extrapolation error, where the learned reward function may fail to explain the task correctly and misguide the agent in unseen…

Machine Learning · Computer Science 2023-02-22 Sheng Yue , Guanbo Wang , Wei Shao , Zhaofeng Zhang , Sen Lin , Ju Ren , Junshan Zhang

In warehouses, specialized agents need to navigate, avoid obstacles and maximize the use of space in the warehouse environment. Due to the unpredictability of these environments, reinforcement learning approaches can be applied to complete…

In a recent work, we proposed a graph-based manifold learning scheme for the nonlinear Galerkin-reduction of quasi-static solid mechanical problems [1]. The resulting nonlinear approximation spaces can closely and flexibly represent…

Computational Engineering, Finance, and Science · Computer Science 2025-09-01 Erik Faust , Lisa Scheunemann

This note concerns a class of matrix Riccati equations associated with stochastic linear-quadratic optimal control problems with indefinite state and control weighting costs. A novel sufficient condition of solvability of such equations is…

Optimization and Control · Mathematics 2013-12-30 Kai Du

We propose an SQP algorithm for mathematical programs with vanishing constraints which solves at each iteration a quadratic program with linear vanishing constraints. The algorithm is based on the newly developed concept of $\mathcal…

Optimization and Control · Mathematics 2016-11-28 Matúš Benko , Helmut Gfrerer

We present two matrix-free methods for approximately solving exact penalty subproblems that arise when solving large-scale optimization problems. The first approach is a novel iterative re-weighting algorithm (IRWA), which iteratively…

Optimization and Control · Mathematics 2017-01-02 James V. Burke , Frank E. Curtis , Hao Wang , Jiashan Wang

A new method, the Dynamical Systems Method (DSM), justified recently, is applied to solving ill-conditioned linear algebraic system (ICLAS). The DSM gives a new approach to solving a wide class of ill-posed problems. In this paper a new…

Numerical Analysis · Mathematics 2009-12-04 Sapto W. Indratno , A. G. Ramm

Quality Diversity (QD) has emerged as a powerful alternative optimization paradigm that aims at generating large and diverse collections of solutions, notably with its flagship algorithm MAP-ELITES (ME) which evolves solutions through…

Neural and Evolutionary Computing · Computer Science 2023-06-16 Thomas Pierrot , Arthur Flajolet

While the Matrix Generalized Inverse Gaussian ($\mathcal{MGIG}$) distribution arises naturally in some settings as a distribution over symmetric positive semi-definite matrices, certain key properties of the distribution and effective ways…

Machine Learning · Statistics 2016-08-23 Farideh Fazayeli , Arindam Banerjee