English
Related papers

Related papers: Time-Uniform Self-Normalized Concentration for Vec…

200 papers

We study rates of convergence in central limit theorems for the partial sum of squares of general Gaussian sequences, using tools from analysis on Wiener space. No assumption of stationarity, asymptotically or otherwise, is made. The main…

Probability · Mathematics 2017-06-09 Soukaina Douissi , Khalifa Es-Sebaiy , Frederi G. Viens

This note describes non-asymptotic variance and tail bounds for order statistics of samples of independent identically distributed random variables. Those bounds are checked to be asymptotically tight when the sampling distribution belongs…

Probability · Mathematics 2012-11-05 Stephane Boucheron , Maud Thomas

Variational inequalities (VIs) are a broad class of optimization problems encompassing machine learning problems ranging from standard convex minimization to more complex scenarios like min-max optimization and computing the equilibria of…

Machine Learning · Computer Science 2025-02-20 Eric Zhao , Tatjana Chavdarova , Michael Jordan

Globally normalized neural sequence models are considered superior to their locally normalized equivalents because they may ameliorate the effects of label bias. However, when considering high-capacity neural parametrizations that condition…

Machine Learning · Computer Science 2019-04-16 Kartik Goyal , Chris Dyer , Taylor Berg-Kirkpatrick

Conditions for geometric ergodicity of multivariate autoregressive conditional heteroskedasticity (ARCH) processes, with the so-called BEKK (Baba, Engle, Kraft, and Kroner) parametrization, are considered. We show for a class of BEKK-ARCH…

Statistics Theory · Mathematics 2017-12-06 Rasmus Pedersen , Olivier Wintenberger

Let $\bX=\{X_n\}_{n\geq 1}$ and $\bY=\{Y_n\}_{n\geq 1}$ be two independent random sequences. We obtain rates of convergence to the normal law of randomly weighted self-normalized sums $$ \psi_n(\bX,\bY)=\sum_{i=1}^nX_iY_i/V_n,\quad…

Probability · Mathematics 2011-09-28 Siegfried Hoermann , Yvik Swan

A generalization of the Bernstein matrix concentration inequality to random tensors of general order is proposed. This generalization is based on the use of Einstein products between tensors, from which a strong link can be established…

Statistics Theory · Mathematics 2021-05-31 Z. Luo , L. Qi , Ph. L. Toint

Transformer models have achieved remarkable results in a wide range of applications. However, their scalability is hampered by the quadratic time and memory complexity of the self-attention mechanism concerning the sequence length. This…

Machine Learning · Computer Science 2024-02-27 Yury Nahshan , Joseph Kampeas , Emir Haleva

A new multivariate integer-valued Generalized AutoRegressive Conditional Heteroscedastic process based on a multivariate Poisson generalized inverse Gaussian distribution is proposed. The estimation of parameters of the proposed…

Computation · Statistics 2023-07-03 Yuhyeong Jang , Raanju R. Sundararajan , Wagner Barreto-Souza

Neural networks have become ubiquitous tools for solving signal and image processing problems, and they often outperform standard approaches. Nevertheless, training neural networks is a challenging task in many applications. The prevalent…

Optimization and Control · Mathematics 2022-10-28 Patrick L. Combettes , Jean-Christophe Pesquet , Audrey Repetti

Large deviation inequalities for ergodic sums is an important subject since the seminal contribution of Bernstein for independent random variables with finite variances, followed by the Chernoff method and the Hoefding result for…

Probability · Mathematics 2025-12-12 Miguel Abadi

In this paper, we present a new framework to obtain tail inequalities for sums of random matrices. Compared with existing works, our tail inequalities have the following characteristics: 1) high feasibility--they can be used to study the…

Machine Learning · Computer Science 2019-10-10 Chao Zhang , Min-Hsiu Hsieh , Dacheng Tao

The aim of this paper is to establish the uniform convergence of the densities of a sequence of random variables, which are functionals of an underlying Gaussian process, to a normal density. Precise estimates for the uniform distance are…

Probability · Mathematics 2013-08-30 Yaozhong Hu , Fei Lu , David Nualart

In this paper, we leverage over-parameterization to design regularization-free algorithms for the high-dimensional single index model and provide theoretical guarantees for the induced implicit regularization phenomenon. Specifically, we…

Machine Learning · Statistics 2021-11-18 Jianqing Fan , Zhuoran Yang , Mengxin Yu

We present moment inequalities for completely degenerate Banach space valued (generalized) U-statistics of arbitrary order. The estimates involve suprema of empirical processes which, in the real-valued case, can be replaced by simpler…

Probability · Mathematics 2007-05-23 Radosław Adamczak

In this paper we study the joint distributional convergence of the largest eigenvalues of the sample covariance matrix of a $p$-dimensional time series with iid entries when $p$ converges to infinity together with the sample size $n$. We…

Probability · Mathematics 2016-08-26 Johannes Heiny , Thomas Mikosch

This paper gives new concentration inequalities for the spectral norm of a wide class of matrix martingales in continuous time. These results extend previously established Freedman and Bernstein inequalities for series of random matrices to…

Probability · Mathematics 2016-10-28 Emmanuel Bacry , Stéphane Gaïffas , Jean-François Muzy

Irreversible aggregation is revisited in view of recent work on renormalization of complex networks. Its scaling laws and phase transitions are related to percolation transitions seen in the latter. We illustrate our points by giving the…

Data Analysis, Statistics and Probability · Physics 2011-08-26 Seung-Woo Son , Golnoosh Bizhani , Claire Christensen , Peter Grassberger , Maya Paczuski

Generalized linear statistics are an unifying class that contains U-statistics, U-quantiles, L-statistics as well as trimmed and winsorized U-statistics. For example, many commonly used estimators of scale fall into this class.…

Statistics Theory · Mathematics 2011-08-19 Martin Wendler

Graph self-supervised learning has sparked a research surge in training informative representations without accessing any labeled data. However, our understanding of graph self-supervised learning remains limited, and the inherent…

Machine Learning · Computer Science 2024-05-17 Taoran Fang , Wei Zhou , Yifei Sun , Kaiqiao Han , Lvbin Ma , Yang Yang