English
Related papers

Related papers: Learning Expected Reward for Switched Linear Contr…

200 papers

We propose $\textit{iterative inversion}$ -- an algorithm for learning an inverse function without input-output pairs, but only with samples from the desired output distribution and access to the forward function. The key challenge is a…

Machine Learning · Computer Science 2023-05-31 Gal Leibovich , Guy Jacob , Or Avner , Gal Novik , Aviv Tamar

We continue our study of the dynamics of mappings with small topological degree on (projective) complex surfaces. Previously, under mild hypotheses, we have constructed an ergodic ``equilibrium'' measure for each such mapping. Here we study…

Dynamical Systems · Mathematics 2009-09-10 Jeffrey Diller , Romain Dujardin , Vincent Guedj

In many applications, it is often necessary to sample the mean value of certain quantity with respect to a probability measure {\mu} on the level set of a smooth function $\xi: \mathbb{R}^d\rightarrow \mathbb{R}^k$, $1\le k < d$. A…

Probability · Mathematics 2019-09-25 Wei Zhang

In this paper, we propose an adaptive event-triggered reinforcement learning control for continuous-time nonlinear systems, subject to bounded uncertainties, characterized by complex interactions. Specifically, the proposed method is…

Machine Learning · Computer Science 2024-10-01 Umer Siddique , Abhinav Sinha , Yongcan Cao

This paper contains two parts. In the first part, we study the ergodicity of periodic measures of random dynamical systems on a separable Banach space. We obtain that the periodic measure of the continuous time skew-product dynamical system…

Probability · Mathematics 2021-03-12 Chunrong Feng , Baoyou Qu , Huaizhong Zhao

We consider reinforcement learning (RL) in episodic Markov decision processes (MDPs) with linear function approximation under drifting environment. Specifically, both the reward and state transition functions can evolve over time but their…

Machine Learning · Computer Science 2024-04-16 Huozhi Zhou , Jinglin Chen , Lav R. Varshney , Ashish Jagmohan

In this article, we study the ergodic risk-sensitive control problem for controlled regime-switching diffusions. Under a blanket stability hypothesis, we solve the associated nonlinear eigenvalue problem for weakly coupled systems and…

Optimization and Control · Mathematics 2022-07-18 Anup Biswas , Somnath Pradhan

We study the problem of system identification for stochastic continuous-time dynamics, based on a single finite-length state trajectory. We present a method for estimating the possibly unstable open-loop matrix by employing properly…

Machine Learning · Statistics 2025-09-30 Reza Sadeghi Hafshejani , Mohamad Kazem Shirani Fradonbeh

In this paper we consider the problem of learning the optimal policy for uncontrolled restless bandit problems. In an uncontrolled restless bandit problem, there is a finite set of arms, each of which when pulled yields a positive reward.…

Optimization and Control · Mathematics 2015-01-30 Cem Tekin , Mingyan Liu

We derive consistency and asymptotic normality results for quasi-maximum likelihood methods for drift parameters of ergodic stochastic processes observed in discrete time in an underlying continuous-time setting. The special feature of our…

Statistics Theory · Mathematics 2021-09-20 Teppei Ogihara , Mitja Stadje

We consider continuous-time random walk models described by arbitrary sojourn time probability density functions. We find a general expression for the distribution of time-averaged observables for such systems, generalizing some recent…

Statistical Mechanics · Physics 2010-09-10 Alberto Saa , Roberto Venegeroles

Via operator theoretic methods, we formalize the concentration phenomenon for a given observable `$r$' of a discrete time Markov chain with `$\mu_{\pi}$' as invariant ergodic measure, possibly having support on an unbounded state space. The…

Machine Learning · Computer Science 2023-06-01 Muhammad Abdullah Naeem , Miroslav Pajic

We present a version of the stochastic maximum principle (SMP) for ergodic control problems. In particular we give necessary (and sufficient) conditions for optimality for controlled dissipative systems in finite dimensions. The strategy we…

Probability · Mathematics 2019-08-05 Carlo Orrieri , Gianmario Tessitore , Petr Veverka

We study learning of probability distributions characterized by an unknown symmetry direction. Based on an entropic performance measure and the variational method of statistical mechanics we develop exact upper and lower bounds on the…

Disordered Systems and Neural Networks · Physics 2009-11-07 D. Herschkowitz , M. Opper

In this work, we study non-asymptotic bounds on correlation between two time realizations of stable linear systems with isotropic Gaussian noise. Consequently, via sampling from a sub-trajectory and using \emph{Talagrands'} inequality, we…

Machine Learning · Statistics 2023-04-05 Muhammad Abdullah Naeem

We prove existence of (at most denumerable many) absolutely continuous invariant probability measures for random one-dimensional dynamical systems with asymptotic expansion. If the rate of expansion (Lyapunov exponents) is bounded away from…

Dynamical Systems · Mathematics 2014-11-18 Vitor Araujo , Javier Solano

Learning-based control methods typically assume stationary system dynamics, an assumption often violated in real-world systems due to drift, wear, or changing operating conditions. We study reinforcement learning for control under…

Machine Learning · Computer Science 2026-04-03 Klemens Iten , Bruce Lee , Chenhao Li , Lenart Treven , Andreas Krause , Bhavya Sukhija

This paper proposes a new adaptation methodology to find the control inputs for a class of nonlinear systems with time-varying bounded uncertainties. The proposed method does not require any prior knowledge of the uncertainties including…

Optimization and Control · Mathematics 2018-03-16 Yi-Wen Liao , Selina Pan , Francesco Borrelli , J. Karl Hedrick

We study multi-objective reinforcement learning with nonlinear preferences over trajectories. That is, we maximize the expected value of a nonlinear function over accumulated rewards (expected scalarized return or ESR) in a multi-objective…

Machine Learning · Computer Science 2025-02-19 Nianli Peng , Muhang Tian , Brandon Fain

For a large class of transitive non-hyperbolic systems, we construct nonhyperbolic ergodic measures with entropy arbitrarily close to its maximal possible value. The systems we consider are partially hyperbolic with one-dimension central…

Dynamical Systems · Mathematics 2022-07-13 Lorenzo J. Díaz , Katrin Gelfert , Michał Rams