English
Related papers

Related papers: Learning Expected Reward for Switched Linear Contr…

200 papers

In this paper, we obtain some preliminary results on stochastic control theory for time-varying linear systems both continuous and discrete, and further apply to aperiod sample-data linear systems. The Ito's lemma is utilized in this…

Systems and Control · Computer Science 2018-02-27 Chunhe Hu , Dan Wu , Junguo Zhang , Zongji Chen

It is well known that ergodic invariant measures for order preserving two-sided time random dynamical systems(RDS) on the real line $\mathbb R$ are Dirac. In the present note this is shown to hold also for one-sided time RDS.

Dynamical Systems · Mathematics 2026-02-18 Hans Crauel

We study ergodic properties of a family of traffic maps acting in the space of bi-infinite sequences of real numbers. The corresponding dynamics mimics the motion of vehicles in a simple traffic flow, which explains the name. Using…

Dynamical Systems · Mathematics 2015-06-11 Michael Blank

We revisit closed-loop performance guarantees for Model Predictive Control in the deterministic and stochastic cases, which extend to novel performance results applicable to receding horizon control of Partially Observable Markov Decision…

Optimization and Control · Mathematics 2020-05-01 Martin A. Sehr , Robert R. Bitmead

Designing a stabilizing controller for nonlinear systems is a challenging task, especially for high-dimensional problems with unknown dynamics. Traditional reinforcement learning algorithms applied to stabilization tasks tend to drive the…

Systems and Control · Electrical Eng. & Systems 2024-09-16 Thanin Quartz , Ruikun Zhou , Hans De Sterck , Jun Liu

This paper studies the adaptive optimal stationary control of continuous-time linear stochastic systems with both additive and multiplicative noises, using reinforcement learning techniques. Based on policy iteration, a novel off-policy…

Systems and Control · Electrical Eng. & Systems 2021-12-07 Bo Pang , Zhong-Ping Jiang

In this article, we pay attention to transitive dynamical systems having the shadowing property and the entropy functions are upper semicontinuous. As for these dynamical systems, when we consider ergodic optimization restricted on the…

Dynamical Systems · Mathematics 2021-12-24 Wanshan Lin , Xueting Tian

We consider stability analysis of constrained switching linear systems in which the dynamics is unknown and whose switching signal is constrained by an automaton. We propose a data-driven Lyapunov framework for providing probabilistic…

Systems and Control · Electrical Eng. & Systems 2022-07-15 Adrien Banse , Zheming Wang , Raphaël M. Jungers

This paper aims to establish an entropy-regularized value-based reinforcement learning method that can ensure the monotonic improvement of policies at each policy update. Unlike previously proposed lower-bounds on policy improvement in…

Machine Learning · Computer Science 2020-08-26 Lingwei Zhu , Takamitsu Matsubara

There has been much recent progress in forecasting the next observation of a linear dynamical system (LDS), which is known as the improper learning, as well as in the estimation of its system matrices, which is known as the proper learning…

Optimization and Control · Mathematics 2024-02-28 Quan Zhou , Jakub Marecek

If $\mathcal{A}$ is a finite set (alphabet), the shift dynamical system consists of the space $\mathcal{A}^{\mathbb{N}}$ of sequences with entries in $\mathcal{A}$, along with the left shift operator $S$. Closed $S$-invariant subsets are…

Dynamical Systems · Mathematics 2020-03-05 Michael Damron , Jon Fickenscher

We consider a large family of discrete and continuous time controlled Markov processes and study an ergodic risk-sensitive minimization problem. Under a blanket stability assumption, we provide a complete analysis to this problem. In…

Optimization and Control · Mathematics 2022-07-18 Anup Biswas , Somnath Pradhan

Multiscale stochastic dynamical systems have been widely adopted to a variety of scientific and engineering problems due to their capability of depicting complex phenomena in many real world applications. This work is devoted to…

Machine Learning · Statistics 2024-01-02 Lingyu Feng , Ting Gao , Min Dai , Jinqiao Duan

In this paper, we present a numerical framework for constructing bounds on stationary performance measures of random walks in the positive orthant using the Markov reward approach. These bounds are established in terms of stationary…

Probability · Mathematics 2018-11-22 Xinwei Bai , Jasper Goseling

This paper focuses on time-varying delayed stochastic differential systems with stochastically switching parameters formulated by a unified switching behavior combining a discrete adapted process and a Cox process. Unlike prior studies…

Dynamical Systems · Mathematics 2024-01-30 Xinyu Wu , Zidong Wang , Wenlian Lu

This work proposes a unified control architecture that couples a Reinforcement Learning (RL)-driven controller with a disturbance-rejection Extended State Observer (ESO), complemented by an Event-Triggered Mechanism (ETM) to limit…

Optimization and Control · Mathematics 2026-01-01 Ningwei Bai , Chi Pui Chan , Qichen Yin , Tengyang Gong , Yunda Yan , Zezhi Tang

We propose a Thompson sampling-based learning algorithm for the Linear Quadratic (LQ) control problem with unknown system parameters. The algorithm is called Thompson sampling with dynamic episodes (TSDE) where two stopping criteria…

Systems and Control · Computer Science 2017-09-14 Yi Ouyang , Mukul Gagrani , Rahul Jain

In this paper, we study concentration phenomena of zero-noise limits of invariant measures for stochastic differential equations defined on $\mathbb{R}^d$ with locally Lipschitz continuous coefficients and more than one ergodic state. Under…

Probability · Mathematics 2022-02-16 Zhao Dong , Fan Gu , Liang Li

In this paper we introduce a new kind of Backward Stochastic Differential Equations, called ergodic BSDEs, which arise naturally in the study of optimal ergodic control. We study the existence, uniqueness and regularity of solution to…

Probability · Mathematics 2007-07-31 Marco Fuhrman , Ying Hu , Gianmario Tessitore

The alignment of Large Language Models (LLMs) is critically dependent on reward models trained on costly human preference data. While recent work explores bypassing this cost with AI feedback, these methods often lack a rigorous theoretical…

Computation and Language · Computer Science 2025-07-01 Yi-Chen Li , Tian Xu , Yang Yu , Xuqin Zhang , Xiong-Hui Chen , Zhongxiang Ling , Ningjing Chao , Lei Yuan , Zhi-Hua Zhou