English
Related papers

Related papers: Large and moderate deviation principles for averag…

200 papers

The stochastic gradient descent (SGD) algorithm has been widely used in statistical estimation for large-scale data due to its computational and memory efficiency. While most existing works focus on the convergence of the objective function…

Machine Learning · Statistics 2023-11-02 Xi Chen , Jason D. Lee , Xin T. Tong , Yichen Zhang

This paper uses a minimum divergence framework to introduce a new way of calculating model weights that can be used to average probabilistic predictions from statistical and machine learning models. The method is general and can be applied…

Machine Learning · Statistics 2026-04-28 Olav Benjamin Vassend

We investigate three types of averaging principles and the normal deviation for multi-scale stochastic differential equations (in short, SDEs) with polynomial nonlinearity. More specifically, we first demonstrate the strong convergence of…

Dynamical Systems · Mathematics 2023-08-22 Mengyu Cheng , Zhenxin Liu , Michael Röckner

We provide a general theorem on the asymptotic behavior of stochastic processes that conform to a relaxed supermartingale condition. The distinguishing feature of our result is that it provides quantitative convergence guarantees at a much…

Optimization and Control · Mathematics 2026-05-11 Morenikeji Neri , Nicholas Pischke , Thomas Powell

The work of Gantert, Kim, and Ramanan [Large deviations for random projections of $\ell^p$ balls, Ann. Probab. 45 (6B), 2017] has initiated and inspired a new direction of research in the asymptotic theory of geometric functional analysis.…

Functional Analysis · Mathematics 2024-03-08 Joscha Prochno

We present here a simple method for computing the large deviation of long time average for stochastic jump processes. We show that the computation of the rate function can be reduced to that of a partial differential equation governing the…

Statistical Mechanics · Physics 2020-04-22 Bahram Houchmandzadeh

We marry ideas from deep neural networks and approximate Bayesian inference to derive a generalised class of deep, directed generative models, endowed with a new algorithm for scalable inference and learning. Our algorithm introduces a…

Machine Learning · Statistics 2014-06-02 Danilo Jimenez Rezende , Shakir Mohamed , Daan Wierstra

We propose a general approach to construct weighted likelihood estimating equations with the aim of obtaining robust parameter estimates. We modify the standard likelihood equations by incorporating a weight that reflects the statistical…

Statistics Theory · Mathematics 2025-07-24 Claudio Agostinelli , Ayanendranath Basu , Giulia Bertagnolli , Arun Kumar Kuchibhotla

We analyze a stochastic approximation algorithm for decision-dependent problems, wherein the data distribution used by the algorithm evolves along the iterate sequence. The primary examples of such problems appear in performative prediction…

Optimization and Control · Mathematics 2024-05-15 Joshua Cutler , Mateo Díaz , Dmitriy Drusvyatskiy

Stochastic Gradient Descent (SGD) is one of the most popular algorithms in statistical and machine learning due to its computational and memory efficiency. Various averaging schemes have been proposed to accelerate the convergence of SGD in…

Machine Learning · Statistics 2025-04-08 Ziyang Wei , Wanrong Zhu , Wei Biao Wu

We propose a new method of estimation in high-dimensional linear regression model. It allows for very weak distributional assumptions including heteroscedasticity, and does not require the knowledge of the variance of random errors. The…

Statistics Theory · Mathematics 2013-04-16 Eric Gautier , Alexandre Tsybakov

Moderate deviation principles for stochastic differential equations driven by a Poisson random measure (PRM) in finite and infinite dimensions are obtained. Proofs are based on a variational representation for expected values of positive…

Probability · Mathematics 2014-01-29 Amarjit Budhiraja , Paul Dupuis , Arnab Ganguly

We introduce data structures for solving robust regression through stochastic gradient descent (SGD) by sampling gradients with probability proportional to their norm, i.e., importance sampling. Although SGD is widely used for large scale…

Machine Learning · Computer Science 2022-07-19 Sepideh Mahabadi , David P. Woodruff , Samson Zhou

This article investigates discrete-time approximations of stochastic integrals driven by semimartingales with jumps via weighted bounded mean oscillation (BMO) approach. This approach enables $L_p$-estimates, $p \in (2, \infty)$, for the…

Probability · Mathematics 2021-12-14 Nguyen Tran Thuan

Subgradient algorithms for training support vector machines have been quite successful for solving large-scale and online learning problems. However, they have been restricted to linear kernels and strongly convex formulations. This paper…

Machine Learning · Computer Science 2011-11-04 Sangkyun Lee , Stephen J. Wright

Cr\'epey, Frikha, and Louzi (2025) introduced a nested stochastic approximation algorithm and its multilevel acceleration to compute the value-at-risk and expected shortfall of a random financial loss. We hereby establish central limit…

Risk Management · Quantitative Finance 2026-04-14 Stéphane Crépey , Noufel Frikha , Azar Louzi , Gilles Pagès

The Alternating Direction Method of Multipliers (ADMM) has been studied for years. The traditional ADMM algorithm needs to compute, at each iteration, an (empirical) expected loss function on all training examples, resulting in a…

Machine Learning · Statistics 2014-06-10 Peilin Zhao , Jinwei Yang , Tong Zhang , Ping Li

Owing to the recent advances in "Big Data" modeling and prediction tasks, variational Bayesian estimation has gained popularity due to their ability to provide exact solutions to approximate posteriors. One key technique for approximate…

Machine Learning · Computer Science 2018-03-01 Hamza Anwar , Quanyan Zhu

Estimating probabilistic deformable template models is a new approach in the fields of computer vision and probabilistic atlases in computational anatomy. A first coherent statistical framework modelling the variability as a hidden random…

Computation · Statistics 2009-01-16 Stéphanie Allassonnière , Estelle Kuhn

This work proposes a machine-learning framework for constructing statistical models of errors incurred by approximate solutions to parameterized systems of nonlinear equations. These approximate solutions may arise from early termination of…

Numerical Analysis · Computer Science 2019-02-18 Brian A. Freno , Kevin T. Carlberg