中文
相关论文

相关论文: Learning Expected Reward for Switched Linear Contr…

200 篇论文

We study optimal stopping of Feller-Markov processes to maximise an undiscounted functional consisting of running and terminal rewards. In a finite-time horizon setting, we extend classical results to unbounded rewards. In infinite horizon,…

最优化与控制 · 数学 2016-07-21 Jan Palczewski , Lukasz Stettner

We study online prediction of bounded stationary ergodic processes. To do so, we consider the setting of prediction of individual sequences and build a deterministic regression tree that performs asymptotically as well as the best…

统计理论 · 数学 2014-05-12 Pierre Gaillard , Paul Baudin

We study in this paper the problem of adaptive trajectory tracking control for a class of nonlinear systems with parametric uncertainties. We propose to use a modular approach, where we first design a robust nonlinear state feedback which…

系统与控制 · 计算机科学 2015-09-28 Mouhacine Benosman , Amir-massoud Farahmand , Meng Xia

Ergodic optimization is the study of extremal values of asymptotic dynamical quantities such as Birkhoff averages or Lyapunov exponents, and of the orbits or invariant measures that attain them. We discuss some results and problems.

动力系统 · 数学 2018-04-24 Jairo Bochi

This work primarily focuses on an operator inference methodology aimed at constructing low-dimensional dynamical models based on a priori hypotheses about their structure, often informed by established physics or expert insights. Stability…

机器学习 · 计算机科学 2024-03-04 Igor Pontes Duff , Pawan Goyal , Peter Benner

Continual reinforcement learning (continual RL) seeks to formalize the notions of lifelong learning and endless adaptation in RL. In particular, the aim of continual RL is to develop RL agents that can maintain a careful balance between…

机器学习 · 计算机科学 2026-05-05 Juan Sebastian Rojas , Chi-Guhn Lee

We study the optimization of ergodic averages for multi-valued dynamical systems, i.e. where points may have multiple different forward orbits. Under upper semi-continuity assumptions, we show that the maximum space average with respect to…

动力系统 · 数学 2025-06-03 Oliver Jenkinson , Xiaoran Li , Yuexin Liao , Yiwei Zhang

We study the problem of controlling linear time-invariant systems with known noisy dynamics and adversarially chosen quadratic losses. We present the first efficient online learning algorithms in this setting that guarantee $O(\sqrt{T})$…

机器学习 · 计算机科学 2018-06-20 Alon Cohen , Avinatan Hassidim , Tomer Koren , Nevena Lazic , Yishay Mansour , Kunal Talwar

We study stochastic approximation procedures for approximately solving a $d$-dimensional linear fixed point equation based on observing a trajectory of length $n$ from an ergodic Markov chain. We first exhibit a non-asymptotic bound of the…

最优化与控制 · 数学 2024-05-14 Wenlong Mou , Ashwin Pananjady , Martin J. Wainwright , Peter L. Bartlett

We present the convergence rates of synchronous and asynchronous Q-learning for average-reward Markov decision processes, where the absence of contraction poses a fundamental challenge. Existing non-asymptotic results overcome this…

机器学习 · 计算机科学 2026-01-30 Zijun Chen , Zaiwei Chen , Nian Si , Shengbo Wang

This paper studies the spread dynamics of a stochastic SIRS epidemic model with nonlinear incidence and varying population size, which is formulated as a piecewise deterministic Markov process. A threshold dynamic determined by the basic…

动力系统 · 数学 2017-10-26 Dan Li , Shengqiang Liu , Jing'an Cui

This paper presents a class of Dynamic Multi-Armed Bandit problems where the reward can be modeled as the noisy output of a time varying linear stochastic dynamic system that satisfies some boundedness constraints. The class allows many…

机器学习 · 计算机科学 2017-10-10 T. W. U. Madhushani , D. H. S. Maithripala , N. E. Leonard

We study the convergence to equilibrium of an underdamped Langevin equation that is controlled by a linear feedback force. Specifically, we are interested in sampling the possibly multimodal invariant probability distribution of a Langevin…

最优化与控制 · 数学 2022-01-12 Tobias Breiten , Carsten Hartmann , Lara Neureither , Upanshu Sharma

We prove ergodicity of a class of infinite measure preserving systems, called skew-products. More precisely, we consider systems of the form \[ {T_f}:{[0, 1) \times \mathbb{R}}\to{[0, 1) \times \mathbb{R}},\quad {T_f(x, t)}:={(T(x),…

动力系统 · 数学 2024-07-11 Fernando Argentieri , Przemysław Berk , Frank Trujillo

In this paper, we study Random Dynamical Systems (RDSs) of homeomorphisms on the circle without a finite orbit. We characterize the topological dynamics of the associated semigroup by identifying the existence of invariant sets which are…

动力系统 · 数学 2025-01-22 Dominique Malicet , Graccyela Salcedo

For a class of linear switched systems in continuous time a controllability condition implies that state feedbacks allow to achieve almost sure stabilization with arbitrary exponential decay rates. This is based on the Multiplicative…

动力系统 · 数学 2019-01-11 Fritz Colonius , Guilherme Mazanti

In this article, a novel adaptive controller is designed for Euler-Lagrangian systems under predefined time-varying state constraints. The proposed controller could achieve this objective without a priori knowledge of system parameters and,…

系统与控制 · 电气工程与系统科学 2024-09-30 Viswa Narayanan Sankaranarayanan , Sumeet Gajanan Satpute , Spandan Roy , George Nikolakopoulos

We present the observation that the process of stochastic model predictive control can be formulated in the framework of iterated function systems. The latter has a rich ergodic theory that can be applied to study the system's long-run…

最优化与控制 · 数学 2022-10-14 Vyacheslav Kungurtsev , Jakub Marecek , Robert Shorten

Nonlinear dynamical systems can be handily described by the associated Koopman operator, whose action evolves every observable of the system forward in time. Learning the Koopman operator and its spectral decomposition from data is enabled…

机器学习 · 计算机科学 2023-11-09 Vladimir Kostic , Karim Lounici , Pietro Novelli , Massimiliano Pontil

We consider ergodic backward stochastic differential equations, in a setting where noise is generated by a countable state uniformly ergodic Markov chain. We show that for Lipschitz drivers such that a comparison theorem holds, these…

概率论 · 数学 2012-07-25 Samuel N. Cohen , Ying Hu