English
Related papers

Related papers: Uniform Convergence with Square-Root Lipschitz Los…

200 papers

We present a computational and statistical approach for fitting isotonic models under convex differentiable loss functions. We offer a recursive partitioning algorithm which provably and efficiently solves isotonic regression under any such…

Methodology · Statistics 2012-10-09 Ronny Luss , Saharon Rosset

The subgradient method is one of the most fundamental algorithmic schemes for nonsmooth optimization. The existing complexity and convergence results for this method are mainly derived for Lipschitz continuous objective functions. In this…

Optimization and Control · Mathematics 2024-11-01 Xiao Li , Lei Zhao , Daoli Zhu , Anthony Man-Cho So

In this paper, we are concerned with differentially private {stochastic gradient descent (SGD)} algorithms in the setting of stochastic convex optimization (SCO). Most of the existing work requires the loss to be Lipschitz continuous and…

Machine Learning · Statistics 2022-03-23 Puyu Wang , Yunwen Lei , Yiming Ying , Hai Zhang

A version of the fundamental mean-square convergence theorem is proved for stochastic differential equations (SDE) which coefficients are allowed to grow polynomially at infinity and which satisfy a one-sided Lipschitz condition. The…

Numerical Analysis · Mathematics 2013-11-26 M. V. Tretyakov , Z. Zhang

This work presents generalized forgetting recursive least squares (GF-RLS), a generalization of recursive least squares (RLS) that encompasses many extensions of RLS as special cases. First, sufficient conditions are presented for the 1)…

Systems and Control · Electrical Eng. & Systems 2024-05-07 Brian Lai , Dennis S. Bernstein

We propose a single time-scale stochastic subgradient method for constrained optimization of a composition of several nonsmooth and nonconvex functions. The functions are assumed to be locally Lipschitz and differentiable in a generalized…

Optimization and Control · Mathematics 2020-12-22 Andrzej Ruszczynski

Iteration complexities for optimizing smooth functions with first-order algorithms are typically stated in terms of a global Lipschitz constant of the gradient, and near-optimal results are then achieved using fixed step sizes. But many…

Optimization and Control · Mathematics 2026-05-19 Curtis Fox , Aaron Mishkin , Sharan Vaswani , Mark Schmidt

We study the task of learning Generalized Linear models (GLMs) in the agnostic model under the Gaussian distribution. We give the first polynomial-time algorithm that achieves a constant-factor approximation for \textit{any} monotone…

Machine Learning · Computer Science 2025-08-05 Nikos Zarifis , Puqian Wang , Ilias Diakonikolas , Jelena Diakonikolas

We extend the classic convergence rate theory for subgradient methods to apply to non-Lipschitz functions. For the deterministic projected subgradient method, we present a global $O(1/\sqrt{T})$ convergence rate for any convex function…

Optimization and Control · Mathematics 2018-02-28 Benjamin Grimmer

We develop a principled approach to obtain exact computer-aided worst-case guarantees on the performance of second-order optimization methods on classes of univariate functions. We first present a generic technique to derive interpolation…

Optimization and Control · Mathematics 2025-07-01 Anne Rubbens , Nizar Bousselmi , Julien M. Hendrickx , François Glineur

We study the problem of estimating the score function of an unknown probability distribution $\rho^*$ from $n$ independent and identically distributed observations in $d$ dimensions. Assuming that $\rho^*$ is subgaussian and has a…

Statistics Theory · Mathematics 2024-06-13 Andre Wibisono , Yihong Wu , Kaylee Yingxi Yang

The concentration of measure phenomenon in Gauss' space states that every $L$-Lipschitz map $f$ on $\mathbb R^n$ satisfies \[ \gamma_{n} \left(\{ x : | f(x) - M_{f} | \geqslant t \} \right) \leqslant 2 e^{ - \frac{t^2}{ 2L^2} }, \quad t>0,…

Probability · Mathematics 2017-06-30 Petros Valettas

We consider nonparametric estimation of the mean and covariance functions for functional/longitudinal data. Strong uniform convergence rates are developed for estimators that are local-linear smoothers. Our results are obtained in a unified…

Statistics Theory · Mathematics 2012-11-12 Yehua Li , Tailen Hsing

Probabilistic learning is increasingly being tackled as an optimization problem, with gradient-based approaches as predominant methods. When modelling multivariate likelihoods, a usual but undesirable outcome is that the learned model fits…

Machine Learning · Computer Science 2020-10-23 Adrián Javaloy , Isabel Valera

We study unconstrained optimization problems of nonsmooth, nonconvex Lipschitz functions, using only noisy pairwise comparisons governed by a known link function. Our goal is to compute a $(\delta,\varepsilon)$-Goldstein stationary point.…

Optimization and Control · Mathematics 2026-02-10 Taha El Bakkali , El Mahdi Chayti , Omar Saadi

We show that H\"older continuity of the gradient is not only a sufficient condition, but also a necessary condition for the existence of a global upper bound on the error of the first-order Taylor approximation. We also relate this global…

Optimization and Control · Mathematics 2020-01-23 Guillaume O. Berger , P. -A. Absil , Raphaël M. Jungers , Yurii Nesterov

Andreas Maurer in the paper "A vector-contraction inequality for Rademacher complexities" extended the contraction inequality for Rademacher averages to Lipschitz functions with vector-valued domains; He did it replacing the Rademacher…

Probability · Mathematics 2021-07-27 Oscar Zatarain-Vera

We study the iteration complexity of Lipschitz convex optimization problems satisfying a general error bound. We show that for this class of problems, subgradient descent with either Polyak stepsizes or decaying stepsizes achieves minimax…

Optimization and Control · Mathematics 2025-12-17 Alex L. Wang

We present a framework for performing efficient regression in general metric spaces. Roughly speaking, our regressor predicts the value at a new point by computing a Lipschitz extension --- the smoothest function consistent with the…

Machine Learning · Computer Science 2017-04-25 Lee-Ad Gottlieb , Aryeh Kontorovich , Robert Krauthgamer

Majorization-minimization algorithms consist of successively minimizing a sequence of upper bounds of the objective function so that along the iterations the objective function decreases. Such a simple principle allows to solve a large…

Optimization and Control · Mathematics 2025-03-04 Ion Necoara , Daniela Lupu