English
Related papers

Related papers: An Output Feedback Q-learning Algorithm for Optima…

200 papers

We present an approach called Q-probing to adapt a pre-trained language model to maximize a task-specific reward function. At a high level, Q-probing sits between heavier approaches such as finetuning and lighter approaches such as few shot…

Machine Learning · Computer Science 2024-06-04 Kenneth Li , Samy Jelassi , Hugh Zhang , Sham Kakade , Martin Wattenberg , David Brandfonbrener

This paper addresses two important estimation problems for linear systems, namely system identification and model-free state estimation. Our focus is on ARMAX models with unknown parameters. We first provide a reinforcement learning…

Systems and Control · Electrical Eng. & Systems 2022-05-10 Minyue Fu

Nonlinear dynamical systems with input delays pose significant challenges for prediction, estimation, and control due to their inherent complexity and the impact of delays on system behavior. Traditional linear control techniques often fail…

Systems and Control · Electrical Eng. & Systems 2025-11-07 Patrik Valábek , Marek Wadinger , Michal Kvasnica , Martin Klaučo

The Koopman operator is a mathematical tool that allows for a linear description of non-linear systems, but working in infinite dimensional spaces. Dynamic Mode Decomposition and Extended Dynamic Mode Decomposition are amongst the most…

Machine Learning · Computer Science 2021-03-26 Francesco Zanini , Alessandro Chiuso

Koopman operators provide a linear framework for data-driven analyses of nonlinear dynamical systems, but their infinite-dimensional nature presents major computational challenges. In this article, we offer an introductory guide to Koopman…

Numerical Analysis · Mathematics 2025-10-28 Matthew J. Colbrook , Zlatko Drmač , Andrew Horning

We propose a neural network-based model for nonlinear dynamics in continuous time that can impose inductive biases on decay rates and/or frequencies. Inductive biases are helpful for training neural networks especially when training data…

Machine Learning · Statistics 2022-12-27 Tomoharu Iwata , Yoshinobu Kawahara

Incorporating nonlinearity into quantum machine learning is essential for learning a complicated input-output mapping. We here propose quantum algorithms for nonlinear regression, where nonlinearity is introduced with feature maps when…

Quantum Physics · Physics 2018-08-30 Dan-Bo Zhang , Shi-Liang Zhu , Z. D. Wang

Q-learning methods represent a commonly used class of algorithms in reinforcement learning: they are generally efficient and simple, and can be combined readily with function approximators for deep reinforcement learning (RL). However, the…

Machine Learning · Computer Science 2019-02-28 Justin Fu , Aviral Kumar , Matthew Soh , Sergey Levine

High fidelity state preparation represents a fundamental challenge in the application of quantum technology. While the majority of optimal control approaches use feedback to improve the controller, the controller itself often does not…

Quantum Physics · Physics 2021-11-22 Ethan N. Evans , Ziyi Wang , Adam G. Frim , Michael R. DeWeese , Evangelos A. Theodorou

We present task-oriented Koopman-based control that utilizes end-to-end reinforcement learning and contrastive encoder to simultaneously learn the Koopman latent embedding, operator, and associated linear controller within an iterative…

Robotics · Computer Science 2023-11-02 Xubo Lyu , Hanyang Hu , Seth Siriya , Ye Pu , Mo Chen

Online optimal control of quadruped robots would enable them to adapt to varying inputs and changing conditions in real time. A common way of achieving this is linear model predictive control (LMPC), where a quadratic programming (QP)…

Robotics · Computer Science 2025-08-13 Chun-Ming Yang , Pranav A. Bhounsule

This paper presents a data-learned linear Koopman embedding of nonlinear networked dynamics and uses it to enable real-time model predictive emergency voltage control in a power network. The approach involves a novel data-driven…

Systems and Control · Electrical Eng. & Systems 2023-10-06 Ramij R. Hossain , Rahmat Adesunkanmi , Ratnesh Kumar

We address the problem of learning a neural Koopman operator model that provides dissipativity guarantees for an unknown nonlinear dynamical system that is known to be dissipative. We propose a two-stage approach. First, we learn an…

Systems and Control · Electrical Eng. & Systems 2025-10-03 Yuezhu Xu , S. Sivaranjani , Vijay Gupta

This is a review of recent research exploring and extending present-day quantum computing capabilities for fusion energy science applications. We begin with a brief tutorial on both ideal and open quantum dynamics, universal quantum…

Quantum Physics · Physics 2023-02-28 I. Joseph , Y. Shi , M. D. Porter , A. R. Castelli , V. I. Geyko , F. R. Graziani , S. B. Libby , J. L. DuBois

The development of machine learning algorithms has been gathering relevance to address the increasing modelling complexity of manufacturing decision-making problems. Reinforcement learning is a methodology with great potential due to the…

Machine Learning · Computer Science 2023-04-18 Miguel Neves , Miguel Vieira , Pedro Neto

One of the most natural approaches to reinforcement learning (RL) with function approximation is value iteration, which inductively generates approximations to the optimal value function by solving a sequence of regression problems. To…

Machine Learning · Computer Science 2024-06-19 Noah Golowich , Ankur Moitra

This paper develops a switching-system interpretation of Q-learning with linear function approximation (LFA) based on the joint spectral radius (JSR). We derive an exact linear switched model for the mean dynamics and relate convergence to…

Machine Learning · Computer Science 2026-05-20 Donghwan Lee , Han-Dong Lim

This paper studies the problem of output regulation for a class of nonlinear systems experiencing matched input disturbances. It is assumed that the disturbance signal is generated by an external autonomous dynamical system. First, we show…

Systems and Control · Electrical Eng. & Systems 2023-09-18 Bart Kieboom , Maria Bartzioka , Matin Jafarian

We study the convergence of $Q$-learning with linear function approximation. Our key contribution is the introduction of a novel multi-Bellman operator that extends the traditional Bellman operator. By exploring the properties of this…

Machine Learning · Computer Science 2023-10-02 Diogo S. Carvalho , Pedro A. Santos , Francisco S. Melo

We exploit the key idea that nonlinear system identification is equivalent to linear identification of the socalled Koopman operator. Instead of considering nonlinear system identification in the state space, we obtain a novel linear…

Systems and Control · Computer Science 2016-08-30 Alexandre Mauroy , Jorge Goncalves
‹ Prev 1 4 5 6 7 8 10 Next ›