English
Related papers

Related papers: Saddlepoint approximation for Student's t-statisti…

200 papers

Sample selection is pervasive in applied economic studies. This paper develops semiparametric selection models that achieve point identification without relying on exclusion restrictions, an assumption long believed necessary for…

Econometrics · Economics 2025-02-11 Dongwoo Kim , Young Jun Lee

Although stochastic approximation learning methods have been widely used in the machine learning literature for over 50 years, formal theoretical analyses of specific machine learning algorithms are less common because stochastic…

Machine Learning · Statistics 2017-04-21 Richard M. Golden

Saddle-point models arise throughout optimization, optimal transport, robust learning, and control. In many applications, the relevant function f(x,y) is convex in x and concave in y, and preserving this geometry is essential for obtaining…

Optimization and Control · Mathematics 2026-05-29 Xavier Warin

In this paper, it is proved a very general well-posedness result for a class of constrained minimization problems.

Optimization and Control · Mathematics 2007-05-23 Biagio Ricceri

We present new stochastic geometry theorems that give bounds on the probability that $m$ random data classes all contain a point in common in their convex hulls. We apply these stochastic separation theorems to obtain bounds on the…

Probability · Mathematics 2019-07-24 Jesús A. De Loera , Thomas A. Hogan

Gradient clipping is a commonly used technique to stabilize the training process of neural networks. A growing body of studies has shown that gradient clipping is a promising technique for dealing with the heavy-tailed behavior that emerged…

Machine Learning · Computer Science 2023-07-26 Shaojie Li , Yong Liu

This study develops a higher-order asymptotic framework for test-time adaptation (TTA) of Batch Normalization (BN) statistics under distribution shift by integrating classical Edgeworth expansion and saddlepoint approximation techniques…

Machine Learning · Statistics 2025-05-23 Masanari Kimura

The Monge-Kantorovich problem is revisited by means of a variant of the saddle-point method without appealing to $c$-conjugates. A new abstract characterization of the optimal plans is obtained in the case where the cost function takes…

Probability · Mathematics 2013-08-02 Christian Léonard

We consider non-convex stochastic optimization using first-order algorithms for which the gradient estimates may have heavy tails. We show that a combination of gradient clipping, momentum, and normalized gradient descent yields convergence…

Machine Learning · Computer Science 2021-11-10 Ashok Cutkosky , Harsh Mehta

Following the same routine as [SSJ20], we continue to present the theoretical analysis for stochastic gradient descent with momentum (SGD with momentum) in this paper. Differently, for SGD with momentum, we demonstrate it is the two…

Machine Learning · Computer Science 2022-09-13 Bin Shi

A recent line of empirical studies has demonstrated that SGD might exhibit a heavy-tailed behavior in practical settings, and the heaviness of the tails might correlate with the overall performance. In this paper, we investigate the…

Machine Learning · Computer Science 2023-10-31 Krunoslav Lehman Pavasovic , Alain Durmus , Umut Simsekli

Models for extreme values are generally derived from limit results, which are meant to be good enough approximations when applied to finite samples. Depending on the speed of convergence of the process underlying the data, these…

Statistics Theory · Mathematics 2019-02-20 Thomas Lugrin , Anthony C. Davison , Jonathan A. Tawn

Despite Adam demonstrating faster empirical convergence than SGD in many applications, much of the existing theory yields guarantees essentially comparable to those of SGD, leaving the empirical performance gap insufficiently explained. In…

Machine Learning · Computer Science 2026-05-19 Ruinan Jin , Yingbin Liang , Shaofeng Zou

In this paper we propose a primal-dual proximal extragradient algorithm to solve the generalized Dantzig selector (GDS) estimation problem, based on a new convex-concave saddle-point (SP) reformulation. Our new formulation makes it possible…

Machine Learning · Statistics 2016-06-03 Sangkyun Lee , Damian Brzyski , Malgorzata Bogdan

In this paper, we propose a general Tikhonov regularized second-order dynamical system with viscous damping, time scaling and extrapolation coefficients for the convex-concave bilinear saddle point problem. By the Lyapunov function…

Optimization and Control · Mathematics 2026-02-02 Bohan Zhang , Xiaojun Zhang

Stochastic transitivity is central for rank aggregation based on pairwise comparison data. The existing models, including the Thurstone, Bradley-Terry (BT), and nonparametric BT models, adopt a strong notion of stochastic transitivity,…

Methodology · Statistics 2025-10-09 Haoran Zhang , Yunxiao Chen

Stochastic Gradient Descent (SGD) and its Ruppert-Polyak averaged variant (ASGD) lie at the heart of modern large-scale learning, yet their theoretical properties in high-dimensional settings are rarely understood. In this paper, we provide…

Machine Learning · Statistics 2025-10-15 Jiaqi Li , Zhipeng Lou , Johannes Schmidt-Hieber , Wei Biao Wu

High-index saddle dynamics (HiSD) is an effective approach for computing saddle points of a prescribed Morse index and constructing solution landscapes for complex nonlinear systems. However, for problems with ill-conditioned Hessians…

Numerical Analysis · Mathematics 2026-05-25 Bingzhang Huang , Hua Su , Lei Zhang , Jin Zhao

We propose in this paper New Q-Newton's method. The update rule is very simple conceptually, for example $x_{n+1}=x_n-w_n$ where $w_n=pr_{A_n,+}(v_n)-pr_{A_n,-}(v_n)$, with $A_n=\nabla ^2f(x_n)+\delta _n||\nabla f(x_n)||^2.Id$ and…

Optimization and Control · Mathematics 2021-09-10 Tuyen Trung Truong , Tat Dat To , Tuan Hang Nguyen , Thu Hang Nguyen , Hoang Phuong Nguyen , Maged Helmy

We deduce in this paper the sufficient conditions for weak convergence of centered and normed deviation of the u-statistics with values in the space of the real valued continuous function defined on some compact metric space. We obtain also…

Statistics Theory · Mathematics 2016-08-12 E. Ostrovsky , L. Sirota
‹ Prev 1 8 9 10 Next ›