English
Related papers

Related papers: Solving the Model Unavailable MARE using Q-Learnin…

200 papers

Autonomous cyber and cyber-physical systems need to perform decision-making, learning, and control in unknown environments. Such decision-making can be sensitive to multiple factors, including modeling errors, changes in costs, and impacts…

Artificial Intelligence · Computer Science 2023-04-05 Abdullah Al Maruf , Luyao Niu , Bhaskar Ramasubramanian , Andrew Clark , Radha Poovendran

Different from most of the previous works, this paper provides a thorough solution to the fundamental problems of linear-quadratic (LQ) control and stabilization for discrete-time mean-field systems under basic assumptions. Firstly, the…

Optimization and Control · Mathematics 2016-11-15 Huanshui Zhang , Qingyuan Qi

Algebraic Riccati equations with indefinite quadratic terms play an important role in applications related to robust controller design. While there are many established approaches to solve these in case of small-scale dense coefficients,…

Numerical Analysis · Mathematics 2023-01-13 Peter Benner , Jan Heiland , Steffen W. R. Werner

A derivative is a financial security whose value is a function of underlying traded assets and market outcomes. Pricing a financial derivative involves setting up a market model, finding a martingale (``fair game") probability measure for…

Quantum Physics · Physics 2022-09-20 Patrick Rebentrost , Alessandro Luongo , Samuel Bosch , Seth Lloyd

This paper presents a discrete-time option pricing model that is rooted in Reinforcement Learning (RL), and more specifically in the famous Q-Learning method of RL. We construct a risk-adjusted Markov Decision Process for a discrete-time…

Computational Finance · Quantitative Finance 2019-09-04 Igor Halperin

Model-free reinforcement learning (RL) algorithms, such as Q-learning, directly parameterize and update value functions or policies without explicitly modeling the environment. They are typically simpler, more flexible to use, and thus more…

Machine Learning · Computer Science 2018-07-11 Chi Jin , Zeyuan Allen-Zhu , Sebastien Bubeck , Michael I. Jordan

Iterative Refinement (IR) is a classical computing technique for obtaining highly precise solutions to linear systems of equations, as well as linear optimization problems. In this paper, motivated by the limited precision of quantum…

Optimization and Control · Mathematics 2023-12-19 Mohammadhossein Mohammadisiahroudi , Brandon Augustino , Pouya Sampourmahani , Tamás Terlaky

Model-free reinforcement learning has been successfully applied to a range of challenging problems, and has recently been extended to handle large neural network policies and value functions. However, the sample complexity of model-free…

Machine Learning · Computer Science 2016-03-03 Shixiang Gu , Timothy Lillicrap , Ilya Sutskever , Sergey Levine

In this paper we explicitly prove that Integrable System solved by Quantum Inverse Scattering Method can be described with the pure algebraic object (Universal R-matrix) and proper algebraic representations. Namely, on the example of the…

High Energy Physics - Theory · Physics 2008-02-03 Alexander Antonov

This paper considers large-scale nonsymmetric continuous-time algebraic Riccati equations (NAREs) that admit low-rank solutions. Low-rank alternating direction implicit (ADI) methods have proven to be an efficient approach for solving…

Numerical Analysis · Mathematics 2026-04-28 Umair Zulfiqar

We focus on the control of unknown Partial Differential Equations (PDEs). The system dynamics is unknown, but we assume we are able to observe its evolution for a given control input, as typical in a Reinforcement Learning framework. We…

Optimization and Control · Mathematics 2023-08-09 Alessandro Alla , Agnese Pacifico , Michele Palladino , Andrea Pesare

We introduce a numerical method for the numerical solution of the so-called Lur'e matrix equations that arise in balancing-related model reduction and linear-quadratic infinite time horizon optimal control. Based on the fact that the set of…

Numerical Analysis · Mathematics 2011-01-07 Federico Poloni , Timo Reis

Continual learning (CL) is crucial for evaluating adaptability in learning solutions to retain knowledge. Our research addresses the challenge of catastrophic forgetting, where models lose proficiency in previously learned tasks as they…

Scientific machine learning is an emerging field that broadly describes the combination of scientific computing and machine learning to address challenges in science and engineering. Within the context of differential equations, this has…

Machine Learning · Computer Science 2026-04-03 Laurens R. Lueg , Victor Alves , Daniel Schicksnus , John R. Kitchin , Carl D. Laird , Lorenz T. Biegler

Large Language Models (LLMs) exhibit strong potential in mathematical reasoning, yet their effectiveness is often limited by a shortage of high-quality queries. This limitation necessitates scaling up computational responses through…

Artificial Intelligence · Computer Science 2025-05-20 Jingyue Gao , Runji Lin , Keming Lu , Bowen Yu , Junyang Lin , Jianyu Chen

Time-dependent wave equations represent an important class of partial differential equations (PDE) for describing wave propagation phenomena, which are often formulated over unbounded domains. Given a compactly supported initial condition,…

Numerical Analysis · Mathematics 2021-07-21 Changjian Xie , Jingrun Chen , Xiantao Li

A promising method for constructing a data-driven output-feedback control law involves the construction of a model-free observer. The Linear Quadratic Regulator (LQR) optimal control policy can then be obtained by both policy-iteration (PI)…

Optimization and Control · Mathematics 2025-09-24 Liquan Lin , Haoyan Lin , Jie Huang

In an episodic Markov Decision Process (MDP) problem, an online algorithm chooses from a set of actions in a sequence of $H$ trials, where $H$ is the episode length, in order to maximize the total payoff of the chosen actions. Q-learning,…

Machine Learning · Computer Science 2019-07-11 Xu Zhu

In this paper, we develop an approach to recursively estimate the quadratic risk for matrix recovery problems regularized with spectral functions. Toward this end, in the spirit of the SURE theory, a key step is to compute the (weak)…

Optimization and Control · Mathematics 2012-11-07 Charles-Alban Deledalle , Samuel Vaiter , Gabriel Peyré , Jalal Fadili , Charles Dossal

We extend the family of problems that may be implemented on an adiabatic quantum optimizer (AQO). When a quadratic optimization problem has at least one set of discrete controls and the constraints are linear, we call this a quadratic…

Quantum Physics · Physics 2014-07-16 Rishabh Chandra , N. Tobias Jacobson , Jonathan E. Moussa , Steven H. Frankel , Sabre Kais
‹ Prev 1 3 4 5 6 7 10 Next ›