English
Related papers

Related papers: Optimal non-asymptotic bound of the Ruppert-Polyak…

200 papers

We derive optimal asymptotic and non-asymptotic lower bounds on the Widom factors for weighted Chebyshev and orthogonal polynomials on compact subsets of the real line. In the Chebyshev case we extend the optimal non-asymptotic lower bound…

Classical Analysis and ODEs · Mathematics 2024-08-22 Gökalp Alpan , Maxim Zinchenko

We address the problem of upper bounding the mean square error of MCMC estimators. Our analysis is nonasymptotic. We first establish a general result valid for essentially all ergodic Markov chains encountered in Bayesian computation and a…

Methodology · Statistics 2013-12-12 Krzysztof Łatuszyński , Błażej Miasojedow , Wojciech Niemiro

This paper presents a Cramer-Rao bound (CRB) for the estimation of parameters confined to an arbitrary set. Unlike existing results that rely on equality or inequality constraints, manifold structures, or the nonsingularity of the Fisher…

Signal Processing · Electrical Eng. & Systems 2026-01-28 Heedong Do , Angel Lozano

Prior work (Klochkov $\&$ Zhivotovskiy, 2021) establishes at most $O\left(\log (n)/n\right)$ excess risk bounds via algorithmic stability for strongly-convex learners with high probability. We show that under the similar common assumptions…

Machine Learning · Computer Science 2025-10-31 Bowei Zhu , Shaojie Li , Mingyang Yi , Yong Liu

Analysis of Stochastic Gradient Descent (SGD) and its variants typically relies on the assumption of uniformly bounded variance, a condition that frequently fails in practical non-convex settings, such as neural network training, as well as…

Machine Learning · Computer Science 2026-04-21 Arda Fazla , Ege C. Kaya , Antesh Upadhyay , Abolfazl Hashemi

We study the least squares estimator in the residual variance estimation context. We show that the mean squared differences of paired observations are asymptotically normally distributed. We further establish that, by regressing the mean…

Statistics Theory · Mathematics 2013-12-12 Tiejun Tong , Yanyuan Ma , Yuedong Wang

The paper deals with asymptotic properties of the adaptive procedure proposed in the author paper, 2007, for estimating an unknown nonparametric regression. %\cite{GaPe1}. We prove that this procedure is asymptotically efficient for a…

Statistics Theory · Mathematics 2010-02-09 Leonid Galtchouk , Serguei Pergamenchtchikov

Among first order optimization methods, Polyak's heavy ball method has long been known to guarantee the asymptotic rate of convergence matching Nesterov's lower bound for functions defined in an infinite-dimensional space. In this paper, we…

Optimization and Control · Mathematics 2023-05-12 V. Ugrinovskii , I. R. Petersen , I. Shames

Random reshuffling techniques are prevalent in large-scale applications, such as training neural networks. While the convergence and acceleration effects of random reshuffling-type methods are fairly well understood in the smooth setting,…

Optimization and Control · Mathematics 2025-07-29 Junwen Qiu , Xiao Li , Andre Milzarek

We consider stochastic optimization problems involving an expected value of a nonlinear function of a base random vector and a conditional expectation of another function depending on the base random vector, a dependent random vector, and…

Optimization and Control · Mathematics 2024-05-20 Andrzej Ruszczyński , Shangzhe Yang

There is widespread sentiment that it is not possible to effectively utilize fast gradient methods (e.g. Nesterov's acceleration, conjugate gradient, heavy ball) for the purposes of stochastic optimization due to their instability and error…

Machine Learning · Statistics 2018-08-02 Prateek Jain , Sham M. Kakade , Rahul Kidambi , Praneeth Netrapalli , Aaron Sidford

We provide novel theoretical results regarding local optima of regularized $M$-estimators, allowing for nonconvexity in both loss and penalty functions. Under restricted strong convexity on the loss and suitable regularity conditions on the…

Statistics Theory · Mathematics 2015-01-05 Po-Ling Loh , Martin J. Wainwright

This paper derives non-asymptotic error bounds for nonlinear stochastic approximation algorithms in the Wasserstein-$p$ distance. To obtain explicit finite-sample guarantees for the last iterate, we develop a coupling argument that compares…

Machine Learning · Computer Science 2026-02-03 Seo Taek Kong , R. Srikant

Nonconvex-nonconcave minimax optimization has gained widespread interest over the last decade. However, most existing works focus on variants of gradient descent-ascent (GDA) algorithms, which are only applicable to smooth nonconvex-concave…

Optimization and Control · Mathematics 2025-01-17 Jiajin Li , Linglingzhi Zhu , Anthony Man-Cho So

We consider a single stage stochastic program without recourse with a strictly convex loss function. We assume a compact decision space and grid it with a finite set of points. In addition, we assume that the decision maker can generate…

Computation · Statistics 2018-11-20 Prateek Jaiswal , Harsha Honnappa , Raghu Pasupathy

We analyze recurrent neural networks with diagonal hidden-to-hidden weight matrices, trained with gradient descent in the supervised learning setting, and prove that gradient descent can achieve optimality \emph{without} massive…

Machine Learning · Computer Science 2024-10-11 Semih Cayci , Atilla Eryilmaz

Stochastic Proximal Gradient (SPG) methods have been widely used for solving optimization problems with a simple (possibly non-smooth) regularizer in machine learning and statistics. However, to the best of our knowledge no non-asymptotic…

Optimization and Control · Mathematics 2019-11-19 Yi Xu , Rong Jin , Tianbao Yang

We consider a finite impulse response system with centered independent sub-Gaussian design covariates and noise components that are not necessarily identically distributed. We derive non-asymptotic near-optimal estimation and prediction…

Statistics Theory · Mathematics 2019-12-02 Boualem Djehiche , Othmane Mazhar , Cristian R. Rojas

Many practical optimization problems lack strong convexity. Fortunately, recent studies have revealed that first-order algorithms also enjoy linear convergences under various weaker regularity conditions. While the relationship among…

Optimization and Control · Mathematics 2026-02-05 Feng-Yi Liao , Lijun Ding , Yang Zheng

The problem of network-constrained averaging is to compute the average of a set of values distributed throughout a graph G using an algorithm that can pass messages only along graph edges. We study this problem in the noisy setting, in…

Distributed, Parallel, and Cluster Computing · Computer Science 2015-06-15 Nima Noorshams , Martin Wainwright
‹ Prev 1 4 5 6 7 8 10 Next ›