English
Related papers

Related papers: Globally Stable Neural Imitation Policies

200 papers

Given a Markov decision process (MDP), we seek to learn representations for a range of policies to facilitate behavior steering at test time. As policies of an MDP are uniquely determined by their occupancy measures, we propose modeling…

Machine Learning · Computer Science 2026-02-02 Beiming Li , Sergio Rozada , Alejandro Ribeiro

Recent advances in learning-based control leverage deep function approximators, such as neural networks, to model the evolution of controlled dynamical systems over time. However, the problem of learning a dynamics model and a stabilizing…

Systems and Control · Electrical Eng. & Systems 2023-04-05 Youngjae Min , Spencer M. Richards , Navid Azizan

Deep Markov models (DMM) are generative models that are scalable and expressive generalization of Markov models for representation, learning, and inference problems. However, the fundamental stochastic stability guarantees of such models…

Machine Learning · Computer Science 2021-11-09 Ján Drgoňa , Sayak Mukherjee , Jiaxin Zhang , Frank Liu , Mahantesh Halappanavar

Neural networks are effective function approximators, but hard to train in the reinforcement learning (RL) context mainly because samples are correlated. For years, scholars have got around this by employing experience replay or an…

Machine Learning · Computer Science 2020-05-06 Budi Kurniawan , Peter Vamplew , Michael Papasimeon , Richard Dazeley , Cameron Foale

The backpropagation algorithm has promoted the rapid development of deep learning, but it relies on a large amount of labeled data and still has a large gap with how humans learn. The human brain can quickly learn various conceptual…

Neural and Evolutionary Computing · Computer Science 2023-04-25 Yiting Dong , Dongcheng Zhao , Yang Li , Yi Zeng

In this study, we propose new global stabilization approaches for a class of polynomial systems in both model-based and data-driven settings. The existing model-based approach guarantees global asymptotic stability of the closed-loop system…

Optimization and Control · Mathematics 2025-05-21 Huayuan Huang , M. Kanat Camlibel , Raffaella Carloni , Henk J. van Waarde

Stochastic gradient descent (SGD) is a powerful optimization technique that is particularly useful in online learning scenarios. Its convergence analysis is relatively well understood under the assumption that the data samples are…

Machine Learning · Computer Science 2024-10-03 Ethan Che , Jing Dong , Xin T. Tong

This paper demonstrates the benefits of imposing stability on data-driven Koopman operators. The data-driven identification of stable Koopman operators (DISKO) is implemented using an algorithm \cite{mamakoukas_stableLDS2020} that computes…

Robotics · Computer Science 2022-03-25 Giorgos Mamakoukas , Ian Abraham , Todd D. Murphey

This paper addresses the problem of model-free reinforcement learning for Robust Markov Decision Process (RMDP) with large state spaces. The goal of the RMDP framework is to find a policy that is robust against the parameter uncertainties…

Machine Learning · Computer Science 2021-02-15 Kishan Panaganti , Dileep Kalathil

Recurrent neural networks (RNNs) are powerful models for processing time-series data, but it remains challenging to understand how they function. Improving this understanding is of substantial interest to both the machine learning and…

Machine Learning · Computer Science 2021-11-03 Jimmy T. H. Smith , Scott W. Linderman , David Sussillo

Spiking Neural Networks (SNNs) offer a promising energy-efficient alternative to Artificial Neural Networks (ANNs) by utilizing sparse and asynchronous processing through discrete spike-based computation. However, the performance of deep…

Neural and Evolutionary Computing · Computer Science 2025-10-10 Eric Jahns , Davi Moreno , Michel A. Kinsy

Infinite-time nonlinear optimal regulation control is widely utilized in aerospace engineering as a systematic method for synthesizing stable controllers. However, conventional methods often rely on linearization hypothesis, while recent…

Systems and Control · Electrical Eng. & Systems 2025-06-13 Han Wang , Di Wu , Lin Cheng , Shengping Gong , Xu Huang

Understanding how the collective activity of neural populations relates to computation and ultimately behavior is a key goal in neuroscience. To this end, statistical methods which describe high-dimensional neural time series in terms of…

Neurons and Cognition · Quantitative Biology 2025-01-14 Amber Hu , David Zoltowski , Aditya Nair , David Anderson , Lea Duncker , Scott Linderman

We study distributed control of networked systems through reinforcement learning, where neural policies must be simultaneously scalable, expressive and stabilizing. We introduce a policy parameterization that embeds Graph Neural Networks…

Systems and Control · Electrical Eng. & Systems 2026-05-27 John Cao , Luca Furieri

Reinforcement learning-based controller design methods often require substantial data in the initial training phase. Moreover, the training process tends to exhibit strong randomness and slow convergence. It often requires considerable time…

Systems and Control · Electrical Eng. & Systems 2025-09-24 Chenxu Ke , Congling Tian , Kaichen Xu , Ye Li , Lingcong Bao

Reinforcement learning (RL) has shown a promising performance in learning optimal policies for a variety of sequential decision-making tasks. However, in many real-world RL problems, besides optimizing the main objectives, the agent is…

Machine Learning · Computer Science 2021-07-30 Ashkan B. Jeddi , Nariman L. Dehghani , Abdollah Shafieezadeh

Policy-gradient methods are widely used in reinforcement learning, yet training often becomes unstable or slows down as learning progresses. We study this phenomenon through the noise-to-signal ratio (NSR) of a policy-gradient estimator,…

Optimization and Control · Mathematics 2026-02-10 Haoyu Han , Heng Yang

We propose Linear Oscillatory State-Space models (LinOSS) for efficiently learning on long sequences. Inspired by cortical dynamics of biological neural networks, we base our proposed LinOSS model on a system of forced harmonic oscillators.…

Machine Learning · Computer Science 2025-06-19 T. Konstantin Rusch , Daniela Rus

State-space models (SSMs) are a highly expressive model class for learning patterns in time series data and for system identification. Deterministic versions of SSMs (e.g. LSTMs) proved extremely successful in modeling complex time series…

Training neural networks to satisfy universal constraints over continuous domains poses unique challenges. Common examples include Lyapunov Neural Networks (Lyapunov NNs) and Physics-Informed Neural Networks (PINNs), where analytical…

Machine Learning · Computer Science 2026-05-12 Siteng Kang , Xinhua Zhang