English
Related papers

Related papers: Random Function Iterations for Consistent Stochast…

200 papers

Entropy regularized algorithms such as Soft Q-learning and Soft Actor-Critic, recently showed state-of-the-art performance on a number of challenging reinforcement learning (RL) tasks. The regularized formulation modifies the standard RL…

Machine Learning · Statistics 2019-10-15 Elena Smirnova , Elvis Dohmatob

This paper is concerned with convergence of stochastic gradient algorithms with momentum terms in the nonconvex setting. A class of stochastic momentum methods, including stochastic gradient descent, heavy ball, and Nesterov's accelerated…

Optimization and Control · Mathematics 2021-10-01 Zixuan Wang , Shanjian Tang

Many applications in networked control require intermittent access of a controller to a system, as in event-triggered systems or information constrained control applications. Motivated by such applications and extending previous work on…

Probability · Mathematics 2015-04-30 Ramiro Zurkowski , Serdar Yüksel , Tamás Linder

Progressive Hedging is a popular decomposition algorithm for solving multi-stage stochastic optimization problems. A computational bottleneck of this algorithm is that all scenario subproblems have to be solved at each iteration. In this…

Distributed, Parallel, and Cluster Computing · Computer Science 2020-09-28 Gilles Bareilles , Yassine Laguel , Dmitry Grishchenko , Franck Iutzeler , Jérôme Malick

The numerical properties of algorithms for finding the intersection of sets depend to some extent on the regularity of the sets, but even more importantly on the regularity of the intersection. The alternating projection algorithm of von…

Optimization and Control · Mathematics 2018-09-24 D. Russell Luke

Pairwise Markov Random Fields (MRFs) or undirected graphical models are parsimonious representations of joint probability distributions. Variables correspond to nodes of a graph, with edges between nodes corresponding to conditional…

Statistics Theory · Mathematics 2018-09-18 Eric Janofsky

A randomised trapezoidal quadrature rule is proposed for continuous functions which enjoys less regularity than commonly required. Indeed, we consider functions in some fractional Sobolev space. Various error bounds for this randomised rule…

Numerical Analysis · Mathematics 2020-12-03 Yue Wu

Based on the idea of randomizing the traditional space theory of functional analysis, random functional analysis has been developed as functional analysis over random metric spaces, random normed modules and random locally convex modules.…

Functional Analysis · Mathematics 2026-03-31 Tiexin Guo , Qiang Tu , Xiaohuan Mu , Yuanyuan Sun

We study the finite-time convergence of TD learning with linear function approximation under Markovian sampling. Existing proofs for this setting either assume a projection step in the algorithm to simplify the analysis, or require a fairly…

Machine Learning · Computer Science 2024-06-27 Aritra Mitra

We consider unconstrained randomized optimization of convex objective functions. We analyze the Random Pursuit algorithm, which iteratively computes an approximate solution to the optimization problem by repeated optimization over a…

Optimization and Control · Mathematics 2012-05-25 Sebastian U. Stich , Christian L. Müller , Bernd Gärtner

One of the proposed solutions to the equilibrium selection problem for agents learning in repeated games is obtained via the notion of stochastic stability. Learning algorithms are perturbed so that the Markov chain underlying the learning…

Computer Science and Game Theory · Computer Science 2012-07-09 John Wicks , Amy Greenwald

The aim of this survey is twofold. First we show how the Markov tower construction is applicable for obtaining finer stochastic properties, like a local limit theorem of probability theory. Here the fundamental method is the study of the…

Dynamical Systems · Mathematics 2007-05-23 Domokos Szász , Tamás Varjú

Iterative Proportional Fitting (IPF), combined with EM, is commonly used as an algorithm for likelihood maximization in undirected graphical models. In this paper, we present two iterative algorithms that generalize upon IPF. The first one…

Machine Learning · Computer Science 2013-01-07 Wim Wiegerinck , Tom Heskes

Recurrent neural networks are widely used in speech and language processing. Due to dependency on the past, standard algorithms for training these models, such as back-propagation through time (BPTT), cannot be efficiently parallelised.…

Audio and Speech Processing · Electrical Eng. & Systems 2021-06-07 Zhengxiong Wang , Anton Ragni

Randomized iterative methods, such as the randomized Kaczmarz method, have gained significant attention for solving large-scale linear systems due to their simplicity and efficiency. Meanwhile, Krylov subspace methods have emerged as a…

Numerical Analysis · Mathematics 2025-05-28 Yonghan Sun , Deren Han , Jiaxin Xie

Motivated by broad applications in reinforcement learning and machine learning, this paper considers the popular stochastic gradient descent (SGD) when the gradients of the underlying objective function are sampled from Markov processes.…

Optimization and Control · Mathematics 2020-04-02 Thinh T. Doan , Lam M. Nguyen , Nhan H. Pham , Justin Romberg

We consider processes which are functions of finite-state Markov chains. It is well known that such processes are rarely Markov. However, such processes are often regular in the following sense: the distant past values of the process have…

Probability · Mathematics 2021-01-05 Steven Berghout , Evgeny Verbitskiy

Existing error-bound-based analyses for stochastic algorithms that exhibit certain descent properties, such as randomized coordinate descent and randomized projection methods, are often limited in scope and typically lead to overly…

Optimization and Control · Mathematics 2026-03-19 Zhichun Yang , Li Jiang , Tianxiang Liu , Man-Chung Yue

We study the almost sure convergence of randomly truncated stochastic algorithms. We present a new convergence theorem which extends the already known results by making vanish the classical condition on the noise terms. The aim of this work…

Probability · Mathematics 2009-06-29 Jérôme Lelong

We present a numerical method to compute expectations of functionals of a piecewise-deterministic Markov process. We discuss time dependent functionals as well as deterministic time horizon problems. Our approach is based on the…

Probability · Mathematics 2012-01-31 Adrien Brandejsky , Benoîte de Saporta , François Dufour
‹ Prev 1 8 9 10 Next ›