English
Related papers

Related papers: Policy iteration using Q-functions: Linear dynamic…

200 papers

Q-functions are widely used in discrete-time learning and control to model future costs arising from a given control policy, when the initial state and input are given. Although some of their properties are understood, Q-functions…

Optimization and Control · Mathematics 2019-02-21 Joseph Warrington

Data informativity provides a theoretical foundation for determining whether collected data are sufficiently informative to achieve specific control objectives in data-driven control frameworks. In this study, we investigate the data…

Optimization and Control · Mathematics 2026-04-22 Taira Kaminaga , Hampei Sasahara

This work focuses on developing a data-driven framework using Koopman operator theory for system identification and linearization of nonlinear systems for control. Our proposed method presents a deep learning framework with recursive…

Systems and Control · Electrical Eng. & Systems 2023-09-11 Madhur Tiwari , George Nehma , Bethany Lusch

Dynamic decision-making under distributional shifts is of fundamental interest in theory and applications of reinforcement learning: The distribution of the environment in which the data is collected can differ from that of the environment…

Machine Learning · Computer Science 2024-09-05 Shengbo Wang , Nian Si , Jose Blanchet , Zhengyuan Zhou

Decision-making under uncertainty is central to many safety-critical applications, where decisions must be guided by probabilistic modeling formalisms. This paper introduces a novel approach to policy synthesis in multi-objective interval…

Systems and Control · Electrical Eng. & Systems 2026-01-08 Negar Monir , Sadegh Soudjani

A dynamic treatment regime is a sequence of decision rules in which each decision rule recommends treatment based on features of patient medical history such as past treatments and outcomes. Existing methods for estimating optimal dynamic…

Methodology · Statistics 2015-05-22 Kristin A. Linn , Eric B. Laber , Leonard A. Stefanski

A stochastic approach to the quantum dynamics randomly modulated in time by a discrete state non-Markovian noise, which possesses an arbitrary non-exponential distribution of the residence times, is developed. The formally exact expression…

Statistical Mechanics · Physics 2007-05-23 Igor Goychuk

This work uses the entropy-regularised relaxed stochastic control perspective as a principled framework for designing reinforcement learning (RL) algorithms. Herein agent interacts with the environment by generating noisy controls…

Machine Learning · Computer Science 2023-09-18 Lukasz Szpruch , Tanut Treetanthiploet , Yufei Zhang

In many applications, it makes sense to solve the least square problems with nonnegative constraints. In this article, we present a new multiplicative iteration that monotonically decreases the value of the nonnegative quadratic programming…

Numerical Analysis · Mathematics 2014-06-05 Xiao Xiao , Donghui Chen

This paper studies adaptive algorithms for simultaneous regulation (i.e., control) and estimation (i.e., learning) of Multiple Input Multiple Output (MIMO) linear dynamical systems. It proposes practical, easy to implement control policies…

Systems and Control · Electrical Eng. & Systems 2020-03-05 Mohamad Kazem Shirani Faradonbeh , Ambuj Tewari , George Michailidis

This paper is concerned with a class of linear-quadratic stochastic large-population problems with partial information, where the individual agent only has access to a noisy observation process related to the state. The dynamics of each…

Optimization and Control · Mathematics 2024-08-20 Min Li , Na Li , Zhen Wu

In this paper new descent line search iterative schemes for unconstrained as well as constrained optimization problems are developed using q-derivative. At every iteration of the scheme, a positive definite matrix is provided which is…

Optimization and Control · Mathematics 2017-02-07 Suvra Kanti Chakraborty , Geetanjali Panda

Reinforcement learning (RL) has achieved significant success across a wide range of domains, however, most existing methods are formulated in discrete time. In this work, we introduce a novel RL method for continuous-time control, where…

Machine Learning · Computer Science 2025-10-21 Chengxiu Hua , Jiawen Gu , Yushun Tang

We study data-driven stabilization of continuous-time systems in autoregressive form when only noisy input-output data are available. First, we provide an operator-based characterization of the set of systems consistent with the data. Next,…

Optimization and Control · Mathematics 2026-02-04 Masashi Wakaiki

A decentralized linear quadratic system with a major agent and a collection of minor agents is considered. The major agent affects the minor agents, but not vice versa. The state of the major agent is observed by all agents. In addition,…

Systems and Control · Electrical Eng. & Systems 2022-07-05 Mohammad Afshari , Aditya Mahajan

This paper addresses the problem of model-free reinforcement learning for Robust Markov Decision Process (RMDP) with large state spaces. The goal of the RMDP framework is to find a policy that is robust against the parameter uncertainties…

Machine Learning · Computer Science 2021-02-15 Kishan Panaganti , Dileep Kalathil

This paper considers the problem of steering the state distribution of a nonlinear stochastic system from an initial Gaussian to a terminal distribution with a specified mean and covariance, subject to probabilistic path constraints. An…

Optimization and Control · Mathematics 2019-09-16 Jack Ridderhof , Kazuhide Okamoto , Panagiotis Tsiotras

We address a learning-based quantum error mitigation method, which utilizes deep neural network applied at the postprocessing stage, and study its performance in presence of different types of quantum noises. We concentrate on the…

Quantum Physics · Physics 2024-02-29 A. A. Zhukov , W. V. Pogosov

We present for the first time an asymptotic convergence analysis of two time-scale stochastic approximation driven by "controlled" Markov noise. In particular, the faster and slower recursions have non-additive controlled Markov noise…

Machine Learning · Computer Science 2020-12-03 Prasenjit Karmakar

Consider the problem of estimating parameters $X^n \in \mathbb{R}^n $, generated by a stationary process, from $m$ response variables $Y^m = AX^n+Z^m$, under the assumption that the distribution of $X^n$ is known. This is the most general…

Information Theory · Computer Science 2017-04-10 Shirin Jalali , Arian Maleki