English
Related papers

Related papers: Concentration Phenomenon for Random Dynamical Syst…

200 papers

Adaptive control of Euler-Lagrange systems is challenging when friction is governed by a finite-horizon internal state that is not directly observable from joint measurements. In this setting, the measured closed-loop state is no longer…

Machine Learning · Computer Science 2026-05-11 Giansalvo Cirrincione , Adriano Fagiolini

The usual random walk on a group (homogeneous both in time and in space) is determined by a probability measure on the group. In a random walk with random transition probabilities this single measure is replaced with a stationary sequence…

Probability · Mathematics 2007-05-23 Vadim A. Kaimanovich , Yuri Kifer , Ben-Zion Rubshtein

This work pioneers regret analysis of risk-sensitive reinforcement learning in partially observable environments with hindsight observation, addressing a gap in theoretical exploration. We introduce a novel formulation that integrates…

Machine Learning · Computer Science 2024-02-29 Tonghe Zhang , Yu Chen , Longbo Huang

In this paper we study a representation problem first considered in a simpler version by Bank and El Karoui [2004]. A key ingredient to this problem is a random measure $\mu$ on the time axis which in the present paper is allowed to have…

Probability · Mathematics 2018-10-22 Peter Bank , David Besslich

We prove a Chernoff-type bound for sums of matrix-valued random variables sampled via a regular (aperiodic and irreducible) finite Markov chain. Specially, consider a random walk on a regular Markov chain and a Hermitian matrix-valued…

Machine Learning · Statistics 2020-10-30 Jiezhong Qiu , Chi Wang , Ben Liao , Richard Peng , Jie Tang

Optimal control in non-stationary Markov decision processes (MDP) is a challenging problem. The aim in such a control problem is to maximize the long-term discounted reward when the transition dynamics or the reward function can change over…

Applications · Statistics 2017-03-03 Taposh Banerjee , Miao Liu , Jonathan P. How

Opacity is a generic security property, that has been defined on (non probabilistic) transition systems and later on Markov chains with labels. For a secret predicate, given as a subset of runs, and a function describing the view of an…

Cryptography and Security · Computer Science 2014-09-02 Béatrice Bérard , Krishnendu Chatterjee , Nathalie Sznajder

This paper presents a convex optimization approach to control the density distribution of autonomous mobile agents with two control modes: ON and OFF. The main new characteristic distinguishing this model from standard Markov decision…

Optimization and Control · Mathematics 2016-12-23 Nazli Demirer , Mahmoud El Chamie , Behcet Acikmese

We consider continuous-time Markov chains on integers which allow transitions to adjacent states only, with alternating rates. We give explicit formulas for probability generating functions, and also for means, variances and state…

Probability · Mathematics 2019-10-30 Luisa Beghin , Claudio Macci , Barbara Martinucci

The first aim of the present note is to quantify the speed of convergence of a conditioned process toward its Q-process under suitable assumptions on the quasi-stationary distribution of the process. Conversely, we prove that, if a…

Probability · Mathematics 2017-04-10 Nicolas Champagnat , Denis Villemonais

We address the reachability problem for continuous-time stochastic dynamic systems. Our objective is to present a unified framework that characterizes the reachable set of a dynamic system in the presence of both stochastic disturbances and…

Systems and Control · Electrical Eng. & Systems 2024-09-04 Saber Jafarpour , Zishun Liu , Yongxin Chen

Inverse optimal control can be used to characterize behavior in sequential decision-making tasks. Most existing work, however, is limited to fully observable or linear systems, or requires the action signals to be known. Here, we introduce…

Machine Learning · Computer Science 2023-10-31 Dominik Straub , Matthias Schultheis , Heinz Koeppl , Constantin A. Rothkopf

A succesful method to describe the asymptotic behavior of a discrete time stochastic process governed by some recursive formula is to relate it to the limit sets of a well chosen mean differential equation. Under an attainability condition,…

Probability · Mathematics 2011-01-19 Mathieu Faure , Gregory Roth

Reinforcement learning in non-stationary environments is challenging due to abrupt and unpredictable changes in dynamics, often causing traditional algorithms to fail to converge. However, in many real-world cases, non-stationarity has some…

Machine Learning · Computer Science 2025-03-25 Mohsen Amiri , Sindri Magnússon

A distributional symmetry is invariance of a distribution under a group of transformations. Exchangeability and stationarity are examples. We explain that a result of ergodic theory provides a law of large numbers: If the group satisfies…

Statistics Theory · Mathematics 2021-11-30 Morgane Austern , Peter Orbanz

We study data-driven learning of robust stochastic control for infinite-horizon systems with potentially continuous state and action spaces. In many managerial settings--supply chains, finance, manufacturing, services, and dynamic…

Machine Learning · Statistics 2025-11-18 Shengbo Wang , Jason Meng , Nian Si , Jose Blanchet , Zhengyuan Zhou

We propose a novel method to directly learn a stochastic transition operator whose repeated application provides generated samples. Traditional undirected graphical models approach this problem indirectly by learning a Markov chain model…

Machine Learning · Statistics 2017-11-08 Anirudh Goyal , Nan Rosemary Ke , Surya Ganguli , Yoshua Bengio

We study the long-term qualitative behavior of randomly perturbed dynamical systems. More specifically, we look at limit cycles of stochastic differential equations (SDE) with Markovian switching, in which the process switches at random…

Probability · Mathematics 2024-07-10 Nguyen H. Du , Alexandru Hening , Dang H. Nguyen , George Yin

Observational entropy -- a quantity that unifies Boltzmann's entropy, Gibbs' entropy, von Neumann's macroscopic entropy, and the diagonal entropy -- has recently been argued to play a key role in a modern formulation of statistical…

Quantum Physics · Physics 2026-03-24 Teruaki Nagasawa , Kohtaro Kato , Eyuri Wakakuwa , Francesco Buscemi

Some general aspects of nonlinear transport phenomena are discussed on the basis of two kinds of formulations obtained by extending Kubo's perturbational scheme of the density matrix and Zubarev's non-equilibrium statistical operator…

Statistical Mechanics · Physics 2011-11-10 Masuo Suzuki
‹ Prev 1 8 9 10 Next ›