English
Related papers

Related papers: Uniform Convergence with Square-Root Lipschitz Los…

200 papers

This work considers the question: what convergence guarantees does the stochastic subgradient method have in the absence of smoothness and convexity? We prove that the stochastic subgradient method, on any semialgebraic locally Lipschitz…

Optimization and Control · Mathematics 2018-05-29 Damek Davis , Dmitriy Drusvyatskiy , Sham Kakade , Jason D. Lee

In this paper we investigate the approximation of continuous functions on the Wasserstein space by smooth functions, with smoothness meant in the sense of Lions differentiability. In particular, in the case of a Lipschitz function we are…

Probability · Mathematics 2023-08-14 Andrea Cosso , Mattia Martini

We show that for convex domains in Euclidean space, Cheeger's isoperimetric inequality, spectral gap of the Neumann Laplacian, exponential concentration of Lipschitz functions, and the a-priori weakest requirement that Lipschitz functions…

Metric Geometry · Mathematics 2008-12-24 Emanuel Milman

We study the optimization of non-convex functions that are not necessarily smooth (gradient and/or Hessian are Lipschitz) using first order methods. Smoothness is a restrictive assumption in machine learning in both theory and practice,…

Optimization and Control · Mathematics 2025-06-27 Daniel Yiming Cao , August Y. Chen , Karthik Sridharan , Benjamin Tang

This paper considers stochastic weakly convex optimization without the standard Lipschitz continuity assumption. Based on new adaptive regularization (stepsize) strategies, we show that a wide class of stochastic algorithms, including the…

Optimization and Control · Mathematics 2024-11-07 Wenzhi Gao , Qi Deng

Generalization bounds which assess the difference between the true risk and the empirical risk, have been studied extensively. However, to obtain bounds, current techniques use strict assumptions such as a uniformly bounded or a Lipschitz…

Machine Learning · Computer Science 2022-11-03 Itai Gat , Yossi Adi , Alexander Schwing , Tamir Hazan

We establish generalization error bounds for stochastic gradient Langevin dynamics (SGLD) with constant learning rate under the assumptions of dissipativity and smoothness, a setting that has received increased attention in the…

Machine Learning · Statistics 2021-11-29 Tyler Farghly , Patrick Rebeschini

We provide an inference procedure for the sharp regression discontinuity design (RDD) under monotonicity, with possibly multiple running variables. Specifically, we consider the case where the true regression function is monotone with…

Econometrics · Economics 2020-12-01 Koohyun Kwon , Soonwoo Kwon

We prove that Picard-Lindel\"of iterations for an arbitrary smooth normal Cauchy problem for PDE converge if we assume a suitable Weissinger-like sufficient condition. This condition includes both a large class of non-analytic PDE or…

Analysis of PDEs · Mathematics 2022-11-03 Paolo Giordano , Lorenzo Luperi Baglini

We derive explicit bounds for the computation of normalizing constants $Z$ for log-concave densities $\pi = \exp(-U)/Z$ with respect to the Lebesgue measure on $\mathbb{R}^d$. Our approach relies on a Gaussian annealing combined with recent…

Methodology · Statistics 2018-03-01 Nicolas Brosse , Alain Durmus , Éric Moulines

We develop an operator-theoretic framework for stability and statistical concentration in nonlinear inverse problems with block-structured parameters. Under a unified set of assumptions combining blockwise Lipschitz geometry, local…

Computer Vision and Pattern Recognition · Computer Science 2026-02-11 Joe-Mei Feng , Hsin-Hsiung Kao

We develop parameter-free algorithms for unconstrained online learning with regret guarantees that scale with the gradient variation $V_T(u) = \sum_{t=2}^T \|\nabla f_t(u)-\nabla f_{t-1}(u)\|^2$. For $L$-smooth convex loss, we provide…

Machine Learning · Computer Science 2026-04-14 Yuheng Zhao , Andrew Jacobsen , Nicolò Cesa-Bianchi , Peng Zhao

Using a perturbation technique, we derive a new approximate filtering and smoothing methodology generalizing along different directions several existing approaches to robust filtering based on the score and the Hessian matrix of the…

Methodology · Statistics 2023-06-06 Giuseppe Buccheri , Giacomo Bormetti , Fulvio Corsi , Fabrizio Lillo

We establish matching upper and lower generalization error bounds for mini-batch Gradient Descent (GD) training with either deterministic or stochastic, data-independent, but otherwise arbitrary batch selection rules. We consider smooth…

Machine Learning · Computer Science 2023-10-24 Konstantinos E. Nikolakakis , Amin Karbasi , Dionysis Kalogerias

It has been observed that certain loss functions can render deep-learning pipelines robust against flaws in the data. In this paper, we support these empirical findings with statistical theory. We especially show that empirical-risk…

Machine Learning · Computer Science 2020-09-15 Johannes Lederer

In many important machine learning applications, the standard assumption of having a globally Lipschitz continuous gradient may fail to hold. This paper delves into a more general $(L_0, L_1)$-smoothness setting, which gains particular…

Optimization and Control · Mathematics 2025-02-07 Chenghan Xie , Chenxi Li , Chuwen Zhang , Qi Deng , Dongdong Ge , Yinyu Ye

This work performs a non-asymptotic analysis of the generalized Lasso under the assumption of sub-exponential data. Our main results continue recent research on the benchmark case of (sub-)Gaussian sample distributions and thereby explore…

Statistics Theory · Mathematics 2023-01-18 Martin Genzel , Christian Kipp

This paper presents uniform estimation and inference theory for a large class of nonparametric partitioning-based M-estimators. The main theoretical results include: (i) uniform consistency for convex and non-convex objective functions;…

Statistics Theory · Mathematics 2025-09-01 Matias D. Cattaneo , Yingjie Feng , Boris Shigida

The classical Gaussian concentration inequality for Lipschitz functions is adapted to a setting where the classical assumptions (i.e. Lipschitz and Gaussian) are not met. The theory is more direct than much of the existing theory designed…

Probability · Mathematics 2022-05-16 Daniel J. Fresen

We study the sequential general online regression, known also as the sequential probability assignments, under logarithmic loss when compared against a broad class of experts. We focus on obtaining tight, often matching, lower and upper…

Machine Learning · Computer Science 2023-02-02 Changlong Wu , Mohsen Heidari , Ananth Grama , Wojciech Szpankowski