English
Related papers

Related papers: Solving the Model Unavailable MARE using Q-Learnin…

200 papers

This paper is concerned with a linear quadratic (LQ, for short) optimal control problem with fixed terminal states and integral quadratic constraints. A Riccati equation with infinite terminal value is introduced, which is uniquely solvable…

Optimization and Control · Mathematics 2017-05-11 Jingrui Sun

We study the autonomous systems of quadratic differential equations of the form $\dot{x}_i(t)=\mathbf{x}(t)^T \mathbf{A}_i \mathbf{x}(t) + \mathbf{v}_i^T \mathbf{x}(t)$ with $\mathbf{x}(t) = (x_1(t),x_2(t),\dots,x_i(t),\dots)$ which, in…

Dynamical Systems · Mathematics 2023-11-22 Ádám Bácsi , Albert Tihamér Kocsis

Mutual exclusion (ME) is one of the most commonly used techniques to handle conflicts in concurrent systems. Traditionally, mutual exclusion algorithms have been designed under the assumption that a process does not fail while…

Distributed, Parallel, and Cluster Computing · Computer Science 2020-08-04 Sahil Dhoked , Neeraj Mittal

Linear-quadratic optimal control problem for systems governed by forward-backward stochastic differential equations has been extensively studied over the past three decades. Recent research has revealed that for forward-backward control…

Optimization and Control · Mathematics 2025-04-22 Qi Lü , Bowen Ma , Hanxiao Wang

The use of target networks is a common practice in deep reinforcement learning for stabilizing the training; however, theoretical understanding of this technique is still limited. In this paper, we study the so-called periodic Q-learning…

Machine Learning · Computer Science 2020-02-25 Donghwan Lee , Niao He

In this paper, a novel Q-learning scheduling method for the current controller of switched reluctance motor (SRM) drive is investigated. Q-learning algorithm is a class of reinforcement learning approaches that can find the best…

Systems and Control · Electrical Eng. & Systems 2020-06-16 Hamad A. Alharkan , Sepehr Saadatmand , Mehdi Ferdowsi , Pourya Shamsi

Can simple algorithms with a good representation solve challenging reinforcement learning problems? In this work, we answer this question in the affirmative, where we take "simple learning algorithm" to be tabular Q-Learning, the "good…

Machine Learning · Computer Science 2020-02-14 Kavosh Asadi , David Abel , Michael L. Littman

I present here a pedagogical introduction to the works by Rashel Tublin and Yan V. Fyodorov on random linear systems with quadratic constraints, using tools from Random Matrix Theory and replicas. These notes illustrate and complement the…

Statistical Mechanics · Physics 2024-03-21 Pierpaolo Vivo

We consider a general statistical linear inverse problem, where the solution is represented via a known (possibly overcomplete) dictionary that allows its sparse representation. We propose two different approaches. A model selection…

Methodology · Statistics 2017-10-31 Felix Abramovich , Daniela De Canditiis , Marianna Pensky

Analytic interpolation problems with rationality and derivative constraints occur in many applications in systems and control. In this paper we present a new method for the multivariable case, which generalizes our previous results on the…

Optimization and Control · Mathematics 2019-03-14 Yufang Cui , Anders Lindquist

The inability to filter out in advance all potentially problematic data from the pre-training of large language models has given rise to the need for methods for unlearning specific pieces of knowledge after training. Existing techniques…

Computation and Language · Computer Science 2026-04-17 Seyun Bae , Seokhan Lee , Eunho Yang

Ensuring safety via safety filters in real-world robotics presents significant challenges, particularly when the system dynamics is complex or unavailable. To handle this issue, learning-based safety filters recently gained popularity,…

Robotics · Computer Science 2024-12-02 Guo Ning Sue , Yogita Choudhary , Richard Desatnik , Carmel Majidi , John Dolan , Guanya Shi

This paper is concerned with stochastic linear quadratic (LQ, for short) optimal control problems in an infinite horizon with constant coefficients. It is proved that the non-emptiness of the admissible control set for all initial state is…

Optimization and Control · Mathematics 2016-10-18 Jingrui Sun , Jiongmin Yong

The worst situation in computing the minimal nonnegative solution of a nonsymmetric algebraic Riccati equation associated with an M-matrix occurs when the corresponding linearizing matrix has two very small eigenvalues, one with positive…

Numerical Analysis · Mathematics 2014-08-26 Bruno Iannazzo , Federico Poloni

Q-learning with function approximation could diverge in the off-policy setting and the target network is a powerful technique to address this issue. In this manuscript, we examine the sample complexity of the associated target Q-learning…

Machine Learning · Computer Science 2022-03-23 Ziniu Li , Tian Xu , Yang Yu

This paper establishes a data-driven solution for infinite horizon linear quadratic Gaussian Mean Field Games with network-coupled heterogeneous agent populations where the dynamics of the agents are unknown. The solution technique relies…

Systems and Control · Electrical Eng. & Systems 2026-02-17 Jean Zhu , Shuang Gao

As it is popular known, Riccati equation is the key basic tool for optimal control in the modern control theory. The solvability conditions of optimal control, stabilization conditions and controller design are all based on the Riccati…

Optimization and Control · Mathematics 2017-12-27 Huanshui Zhang , Juanjuan Xu

The $Q$-learning algorithm is a simple and widely-used stochastic approximation scheme for reinforcement learning, but the basic protocol can exhibit instability in conjunction with function approximation. Such instability can be observed…

Machine Learning · Computer Science 2022-06-03 Andrea Zanette , Martin J. Wainwright

We review a family of algorithms for Lyapunov- and Riccati-type equations which are all related to each other by the idea of \emph{doubling}: they construct the iterate $Q_k = X_{2^k}$ of another naturally-arising fixed-point iteration…

Numerical Analysis · Mathematics 2020-05-19 Federico Poloni

In recent years there has been a collective research effort to find new formulations of reinforcement learning that are simultaneously more efficient and more amenable to analysis. This paper concerns one approach that builds on the linear…

Optimization and Control · Mathematics 2022-10-19 Fan Lu , Prashant Mehta , Sean Meyn , Gergely Neu