English
Related papers

Related papers: Optimal and instance-dependent guarantees for Mark…

200 papers

Two-time-scale stochastic approximation (SA) is an algorithm with coupled iterations which has found broad applications in reinforcement learning, optimization and game control. In this work, we derive mean squared error bounds for…

Machine Learning · Computer Science 2026-02-24 Siddharth Chandak

We consider the problem of approximating the reachability probabilities in Markov decision processes (MDP) with uncountable (continuous) state and action spaces. While there are algorithms that, for special classes of such MDP, provide a…

Systems and Control · Electrical Eng. & Systems 2022-07-13 Kush Grover , Jan Křetínský , Tobias Meggendorfer , Maximilian Weininger

This paper studies fixed step-size stochastic approximation (SA) schemes, including stochastic gradient schemes, in a Riemannian framework. It is motivated by several applications, where geodesics can be computed explicitly, and their use…

Machine Learning · Statistics 2021-02-22 Alain Durmus , Pablo Jiménez , Éric Moulines , Salem Said

The paper concerns the $d$-dimensional stochastic approximation recursion, $$ \theta_{n+1}= \theta_n + \alpha_{n + 1} f(\theta_n, \Phi_{n+1}) $$ where $ \{ \Phi_n \}$ is a stochastic process on a general state space, satisfying a…

Statistics Theory · Mathematics 2024-11-18 Vivek Borkar , Shuhang Chen , Adithya Devraj , Ioannis Kontoyiannis , Sean Meyn

We give a complete characterization of the sampling complexity of best Markovian arm identification in one-parameter Markovian bandit models. We derive instance specific nonasymptotic and asymptotic lower bounds which generalize those of…

Statistics Theory · Mathematics 2020-07-29 Vrettos Moulos

Multi-time-scale stochastic approximation is an iterative algorithm for finding the fixed point of a set of $N$ coupled operators given their noisy samples. It has been observed that due to the coupling between the decision variables and…

Optimization and Control · Mathematics 2024-09-13 Sihan Zeng , Thinh T. Doan

We study the finite-time convergence of TD learning with linear function approximation under Markovian sampling. Existing proofs for this setting either assume a projection step in the algorithm to simplify the analysis, or require a fairly…

Machine Learning · Computer Science 2024-06-27 Aritra Mitra

We propose and study an asymptotically optimal Monte Carlo estimator for steady-state expectations of a d-dimensional reflected Brownian motion. Our estimator is asymptotically optimal in the sense that it requires $\tilde{O}(d)$ (up to…

Probability · Mathematics 2020-01-29 Jose Blanchet , Xinyun Chen , Peter Glynn , Nian Si

This is a technical report that extends and clarifies the results presented in [1]. The model identification problem for asymptotically stable linear time invariant systems is considered. The system output is affected by an additive noise…

Optimization and Control · Mathematics 2018-09-05 Marco Lauricella , Lorenzo Fagiano

We propose a method for approximating solutions to optimization problems involving the global stability properties of parameter-dependent continuous-time autonomous dynamical systems. The method relies on an approximation of the…

Optimization and Control · Mathematics 2013-08-12 Péter Koltai , Alexander Volf

Two-timescale stochastic approximation (TTSA) is among the most general frameworks for iterative stochastic algorithms. This includes well-known stochastic optimization methods such as SGD variants and those designed for bilevel or minimax…

Machine Learning · Statistics 2024-02-15 Jie Hu , Vishwaraj Doshi , Do Young Eun

Regularly varying stochastic processes model extreme dependence between process values at different locations and/or time points. For such processes we propose a two-step parameter estimation of the extremogram, when some part of the domain…

Statistics Theory · Mathematics 2018-08-28 Sven Buhl , Claudia Klüppelberg

We study a decentralized variant of stochastic approximation, a data-driven approach for finding the root of an operator under noisy measurements. A network of agents, each with its own operator and data observations, cooperatively find the…

Machine Learning · Computer Science 2022-06-17 Sihan Zeng , Thinh T. Doan , Justin Romberg

In this work, we present a generalized methodology for analyzing the convergence of quasi-optimal Taylor and Legendre approximations, applicable to a wide class of parameterized elliptic PDEs with finite-dimensional deterministic and…

Analysis of PDEs · Mathematics 2015-08-11 Hoang Tran , Clayton G. Webster , Guannan Zhang

We define a stochastic variant of the proximal point algorithm in the general setting of nonlinear (separable) Hadamard spaces for approximating zeros of the mean of a stochastically perturbed monotone vector field and prove its convergence…

Optimization and Control · Mathematics 2025-10-14 Nicholas Pischke

We study oracle complexity of gradient based methods for stochastic approximation problems. Though in many settings optimal algorithms and tight lower bounds are known for such problems, these optimal algorithms do not achieve the best…

Optimization and Control · Mathematics 2022-06-20 Jingzhao Zhang , Hongzhou Lin , Subhro Das , Suvrit Sra , Ali Jadbabaie

We consider approximate dynamic programming for the infinite-horizon stationary $\gamma$-discounted optimal control problem formalized by Markov Decision Processes. While in the exact case it is known that there always exists an optimal…

Optimization and Control · Mathematics 2013-04-23 Boris Lesner , Bruno Scherrer

In this paper, we study the problem of estimating a Markov chain $X$(signal) from its noisy partial information $Y$, when the transition probability kernel depends on some unknown parameters. Our goal is to compute the conditional…

Probability · Mathematics 2007-05-23 Anastasia Papavasiliou

We investigate the robustness of nonlinear filtering for continuous time finite state Markov chains, observed in white noise, with respect to misspecification of the model parameters. It is shown that the distance between the optimal filter…

Probability · Mathematics 2007-05-23 Pavel Chigansky , Ramon van Handel

We revisit the development of grid based recursive approximate filtering of general Markov processes in discrete time, partially observed in conditionally Gaussian noise. The grid based filters considered rely on two types of state…

Statistics Theory · Mathematics 2016-08-24 Dionysios S. Kalogerias , Athina P. Petropulu
‹ Prev 1 3 4 5 6 7 10 Next ›