English
Related papers

Related papers: Convergence Rate in Nonlinear Two-Time-Scale Stoch…

200 papers

We study the policy evaluation problem in multi-agent reinforcement learning, modeled by a Markov decision process. In this problem, the agents operate in a common environment under a fixed control policy, working together to discover the…

Optimization and Control · Mathematics 2020-01-13 Thinh T. Doan , Siva Theja Maguluri , Justin Romberg

The influence of multiplicative white noise on the resonance capture of strongly nonlinear oscillatory systems under chirped-frequency excitations is investigated. It is assumed that the intensity of the perturbation decays polynomially…

Dynamical Systems · Mathematics 2024-03-22 Oskar A. Sultanov

We consider the nonparametric estimation problem of time-dependent multivariate functions observed in a presence of additive cylindrical Gaussian white noise of a small intensity. We derive minimax lower bounds for the $L^2$-risk in the…

Statistics Theory · Mathematics 2012-11-02 Jérémie Bigot , Theofanis Sapatinas

This paper addresses the gradient flow -- the continuous-time representation of the gradient method -- with the smooth approximation of a non-differentiable objective function and presents convergence analysis framework. Similar to the…

Optimization and Control · Mathematics 2023-12-08 Mitsuru Toyoda , Akatsuki Nishioka , Mirai Tanaka

Despite its popularity in the reinforcement learning community, a provably convergent policy gradient method for continuous space-time control problems with nonlinear state dynamics has been elusive. This paper proposes proximal gradient…

Optimization and Control · Mathematics 2022-12-27 Christoph Reisinger , Wolfgang Stockinger , Yufei Zhang

For complex nonlinear systems, it is challenging to design algorithms that are fast, scalable, and give an accurate approximation of the stability region. This paper proposes a sampling-based approach to address these challenges. By…

Systems and Control · Electrical Eng. & Systems 2024-05-24 Péter Antal , Tamás Péni , Roland Tóth

It is known that state-dependent, multi-step Lyapunov bounds lead to greatly simplified verification theorems for stability for large classes of Markov chain models. This is one component of the "fluid model" approach to stability of…

Optimization and Control · Mathematics 2012-05-18 Serdar Yüksel , Sean P. Meyn

This paper investigates the asymptotic behavior of stochastic recursive inclusions in the presence of non-zero, non-diminishing bias, a setting that frequently arises in zeroth-order optimization, stochastic approximation with…

Optimization and Control · Mathematics 2026-01-19 Anik Kumar Paul , Karthik Shenoy , Arun D. Mahindrakar

In large-scale learning algorithms, the momentum term is usually included in the stochastic sub-gradient method to improve the learning speed because it can navigate ravines efficiently to reach a local minimum. However, step-size and…

Machine Learning · Computer Science 2024-08-07 Wen-Liang Hwang

This paper develops a new approach to the estimation of the degree of boundedness or stability of multidimensional nonlinear systems with time-dependent nonperiodic coefficients-an essential task in various engineering and natural science…

Dynamical Systems · Mathematics 2022-06-16 Mark A. Pinsky

We claim that looking at probability distributions of \emph{finite time} largest Lyapunov exponents, and more precisely studying their large deviation properties, yields an extremely powerful technique to get quantitative estimates of…

Chaotic Dynamics · Physics 2009-10-20 Roberto Artuso , Cesar Manchein

We present a stability analysis framework for the general class of discrete-time linear switching systems for which the switching sequences belong to a regular language. They admit arbitrary switching systems as special cases. Using recent…

Dynamical Systems · Mathematics 2014-11-17 Matthew Philippe , Raphaël M. Jungers

This article establishes the existence of Lyapunov functions for analyzing the stability of a class of state-constrained systems, and it describes algorithms for their numerical computation. The system model consists of a differential…

Optimization and Control · Mathematics 2021-04-14 Marianne Souaiby , Aneel Tanwani , Didier Henrion

This paper investigates the weighted-averaging dynamic for unconstrained and constrained consensus problems. Through the use of a suitably defined adjoint dynamic, quadratic Lyapunov comparison functions are constructed to analyze the…

Optimization and Control · Mathematics 2014-07-30 Angelia Nedich , Ji Liu

This paper is a comprehensive study of a long observed phenomenon of increase in the stability margin and so the rate of convergence of a class of linear systems due to time delay. We use Lambert W function to determine (a) in what systems…

Multiagent Systems · Computer Science 2019-07-23 Hossein Moradian , Solmaz S. Kia

This paper investigates the problem of consensus tracking control of discrete time multi-agent systems under binary-valued communication. Different from most existing studies on consensus tracking, the transmitted information between agents…

Multiagent Systems · Computer Science 2025-03-21 Ting Wang , Zhuangzhuang Qiu , Xiaodong Lu , Yanlong Zhao

On the basis of a local-projective (LP) approach we develop a method of noise reduction in time series that makes use of nonlinear constraints appearing due to the deterministic character of the underlying dynamical system. The Delaunay…

Statistical Mechanics · Physics 2007-05-23 Krzysztof Urbanowicz , Janusz A. Holyst , Thomas Stemler , Hartmut Benner

We consider spatially extended conductance based neuronal models with noise described by a stochastic reaction diffusion equation with additive noise coupled to a control variable with multiplicative noise but no diffusion. We only assume a…

Probability · Mathematics 2020-01-16 Martin Sauer , Wilhelm Stannat

We study the popular distributed consensus method over networks composed of a number of densely connected clusters with a sparse connection between them. In these cluster networks, the method often constitutes two-time-scale dynamics, where…

Optimization and Control · Mathematics 2022-09-14 Amit Dutta , Almuatazbellah M. Boker , Thinh T. Doan

The analysis in Part I revealed interesting properties for subgradient learning algorithms in the context of stochastic optimization when gradient noise is present. These algorithms are used when the risk functions are non-smooth and…

Optimization and Control · Mathematics 2017-04-21 Bicheng Ying , Ali H. Sayed