English
Related papers

Related papers: Learning Expected Reward for Switched Linear Contr…

200 papers

We consider the problem of learning the dynamics of autonomous linear systems (i.e., systems that are not affected by external control inputs) from observations of multiple trajectories of those systems, with finite sample guarantees.…

Systems and Control · Electrical Eng. & Systems 2022-09-27 Lei Xin , George Chiu , Shreyas Sundaram

Multi-agent learning is a promising method to simulate aggregate competitive behaviour in finance. Learning expert agents' reward functions through their external demonstrations is hence particularly relevant for subsequent design of…

Machine Learning · Computer Science 2019-06-13 Jacobo Roa-Vicens , Cyrine Chtourou , Angelos Filos , Francisco Rullan , Yarin Gal , Ricardo Silva

This paper proposes a novel online data-driven adaptive control for unknown linear time-varying systems. Initialized with an empirical feedback gain, the algorithm periodically updates this gain based on the data collected over a short time…

Systems and Control · Electrical Eng. & Systems 2024-01-31 Shenyu Liu , Kaiwen Chen , Jaap Eising

The dynamics of the solutions to a class of conservative SPDEs are analysed from two perspectives: Firstly, a probabilistic construction of a corresponding random dynamical system is given for the first time. Secondly, the existence and…

Probability · Mathematics 2022-06-30 Benjamin Fehrman , Benjamin Gess , Rishabh S. Gvalani

In this article, we introduce a system of stochastic differential equations (SDEs) consisting of time-dependent covariates and consider both fixed and random effects set-ups. We also allow the functional part associated with the drift…

Statistics Theory · Mathematics 2017-10-16 Trisha Maitra , Sourabh Bhattacharya

Sufficient conditions characterizing the asymptotic stability and the hybrid $L_1/\ell_1$-gain of linear positive impulsive systems under minimum and range dwell-time constraints are obtained. These conditions are stated as…

Optimization and Control · Mathematics 2018-10-16 Corentin Briat

We study the limit behaviour of upper and lower bounds on expected time averages in imprecise Markov chains; a generalised type of Markov chain where the local dynamics, traditionally characterised by transition probabilities, are now…

Probability · Mathematics 2021-02-10 Natan T'Joens , Jasper De Bock

The classical Birkhoff ergodic theorem states that for an ergodic Markov process the limiting behaviour of the time average of a function (having finite $p$-th moment, $p\ge1$, with respect to the invariant measure) along the trajectories…

Probability · Mathematics 2017-04-13 Nikola Sandrić

This letter proposes a learning-based bounded synthesis for a semi-Markov decision process (SMDP) with a linear temporal logic (LTL) specification. In the product of the SMDP and the deterministic $K$-co-B\"uchi automaton (d$K$cBA)…

Systems and Control · Electrical Eng. & Systems 2022-04-12 Ryohei Oura , Toshimitsu Ushio

This article is concerned with stability analysis and stabilization of randomly switched nonlinear systems. These systems may be regarded as piecewise deterministic stochastic systems: the discrete switches are triggered by a stochastic…

Optimization and Control · Mathematics 2010-09-08 Debasish Chatterjee , Daniel Liberzon

We derive an asymptotic log-Harnack inequality for nonlinear monotone SPDE driven by possibly degenerate multiplicative noise. Our main tool is the asymptotic coupling by the change of measure. As an application, we show that, under certain…

Probability · Mathematics 2024-09-19 Zhihui Liu

We study a multi-armed bandit problem where the rewards exhibit regime switching. Specifically, the distributions of the random rewards generated from all arms are modulated by a common underlying state modeled as a finite-state Markov…

Machine Learning · Computer Science 2021-02-02 Xiang Zhou , Yi Xiong , Ningyuan Chen , Xuefeng Gao

This paper studies an attitude control system design based on modified Rodrigues parameters feedback. It employs a linear continuous sliding mode controller. The sliding mode controller is able to bring the existence of the sliding motion…

Systems and Control · Electrical Eng. & Systems 2021-02-04 Harry Septanto , Djoko Suprijanto

This work is devoted to the almost sure stabilization of adaptive control systems that involve an unknown Markov chain. The control system displays continuous dynamics represented by differential equations and discrete events given by a…

Probability · Mathematics 2008-07-10 Bernard Bercu , Francois Dufour , G. George Yin

We propose a method for learning dynamical systems from high-dimensional empirical data that combines variational autoencoders and (spatio-)temporal attention within a framework designed to enforce certain scientifically-motivated…

Machine Learning · Computer Science 2023-06-22 Kai Lagemann , Christian Lagemann , Sach Mukherjee

Inverse Reinforcement Learning (IRL) describes the problem of learning an unknown reward function of a Markov Decision Process (MDP) from observed behavior of an agent. Since the agent's behavior originates in its policy and MDP policies…

Artificial Intelligence · Computer Science 2016-04-14 Michael Herman , Tobias Gindele , Jörg Wagner , Felix Schmitt , Wolfram Burgard

In this paper we study the ergodic theory of a class of symbolic dynamical systems $(\O, T, \mu)$ where $T:{\O}\to \O$ the left shift transformation on $\O=\prod_0^\infty\{0,1\}$ and $\mu$ is a $\s$-finite $T$-invariant measure having the…

Dynamical Systems · Mathematics 2007-05-23 Stefano Isola

An adaptive controller with bounded l2-gain from disturbances to errors is derived for linear time-invariant systems with uncertain parameters restricted to a finite set. The gain bound refers to the closed loop system, including the…

Optimization and Control · Mathematics 2024-04-09 Anders Rantzer

We present a data-driven framework for strategy synthesis for partially-known switched stochastic systems. The properties of the system are specified using linear temporal logic (LTL) over finite traces (LTLf), which is as expressive as LTL…

Systems and Control · Electrical Eng. & Systems 2022-03-10 John Jackson , Luca Laurenti , Eric Frew , Morteza Lahijanian

This paper studies the optimal output-feedback control of a linear time-invariant system where a stochastic event-based scheduler triggers the communication between the sensor and the controller. The primary goal of the use of this type of…

Systems and Control · Computer Science 2017-08-10 Burak Demirel , Alex S. Leong , Vijay Gupta , Daniel E. Quevedo