English
Related papers

Related papers: Uniform Convergence with Square-Root Lipschitz Los…

200 papers

While momentum-based accelerated variants of stochastic gradient descent (SGD) are widely used when training machine learning models, there is little theoretical understanding on the generalization error of such methods. In this work, we…

Machine Learning · Computer Science 2024-01-17 Ali Ramezani-Kebrya , Kimon Antonakopoulos , Volkan Cevher , Ashish Khisti , Ben Liang

Operator learning, the approximation of mappings between infinite-dimensional function spaces using machine learning, has gained increasing research attention in recent years. Approximate operators, learned from data, can serve as efficient…

Machine Learning · Computer Science 2025-06-27 Ben Adcock , Michael Griebel , Gregor Maier

In this paper, we propose a coupled tensor norm regularization that could enable the model output feature and the data input to lie in a low-dimensional manifold, which helps us to reduce overfitting. We show this regularization term is…

Optimization and Control · Mathematics 2023-02-24 Ying Gao , Yunfei Qu , Chunfeng Cui , Deren Han

We revisit the geometrically decaying step size given a positive inverse condition number, under which a locally Lipschitz function shows linear convergence. The positivity does not require the function to satisfy convexity, weak convexity,…

Optimization and Control · Mathematics 2025-12-04 Jihun Kim

The contraction inequality for Rademacher averages is extended to Lipschitz functions with vector-valued domains, and it is also shown that in the bounding expression the Rademacher variables can be replaced by arbitrary iid symmetric and…

Machine Learning · Computer Science 2016-05-04 Andreas Maurer

Algorithm- and data-dependent generalization bounds are required to explain the generalization behavior of modern machine learning algorithms. In this context, there exists information theoretic generalization bounds that involve (various…

Machine Learning · Statistics 2023-07-07 Sarah Sachs , Tim van Erven , Liam Hodgkinson , Rajiv Khanna , Umut Simsekli

Diffusion models have made rapid progress in generating high-quality samples across various domains. However, a theoretical understanding of the Lipschitz continuity and second momentum properties of the diffusion process is still lacking.…

Machine Learning · Computer Science 2024-10-15 Yingyu Liang , Zhenmei Shi , Zhao Song , Yufa Zhou

Bubeck and Sellke (2021) pose as an open problem the connection between the law of robustness and robust generalization. The law of robustness states that overparameterization is necessary for models to interpolate robustly; in particular,…

Machine Learning · Computer Science 2026-02-26 Himadri Mandal , Vishnu Varadarajan , Jaee Ponde , Aritra Das , Mihir More , Debayan Gupta

While it is well known that the restricted isometry property (RIP) guarantees uniform sparse recovery from noisy linear measurements, uniform recovery of structured signals from nonlinear observations remains much less understood. This…

Information Theory · Computer Science 2026-04-23 Pedro Abdalla , Radu Balan , Junren Chen

Generalised Bayesian inference updates prior beliefs using a loss function, rather than a likelihood, and can therefore be used to confer robustness against possible mis-specification of the likelihood. Here we consider generalised Bayesian…

Methodology · Statistics 2022-01-12 Takuo Matsubara , Jeremias Knoblauch , François-Xavier Briol , Chris. J. Oates

Distributionally robust optimization has emerged as an attractive way to train robust machine learning models, capturing data uncertainty and distribution shifts. Recent statistical analyses have proved that generalization guarantees of…

Optimization and Control · Mathematics 2025-01-28 Tam Le , Jérôme Malick

We study the problem of global optimization, where we analyze the performance of the Piyavskii--Shubert algorithm and its variants. For any given time duration $T$, instead of the extensively studied simple regret (which is the difference…

Machine Learning · Computer Science 2023-12-29 Kaan Gokcesu , Hakan Gokcesu

We consider the long-term dynamics of the vanishing stepsize subgradient method in the case when the objective function is neither smooth nor convex. We assume that this function is locally Lipschitz and path differentiable, i.e., admits a…

Optimization and Control · Mathematics 2020-06-02 Jerome Bolte , Edouard Pauwels , Rodolfo Rios-Zertuche

One fundamental goal in any learning algorithm is to mitigate its risk for overfitting. Mathematically, this requires that the learning algorithm enjoys a small generalization risk, which is defined either in expectation or in probability.…

Machine Learning · Computer Science 2016-10-04 Ibrahim Alabdulmohsin

In this paper, we consider unregularized online learning algorithms in a Reproducing Kernel Hilbert Spaces (RKHS). Firstly, we derive explicit convergence rates of the unregularized online learning algorithms for classification associated…

Machine Learning · Computer Science 2015-04-28 Yiming Ying , Ding-Xuan Zhou

Projection-free optimization via different variants of the Frank-Wolfe (FW) method has become one of the cornerstones in large scale optimization for machine learning and computational statistics. Numerous applications within these fields…

Optimization and Control · Mathematics 2021-08-03 Pavel Dvurechensky , Kamil Safin , Shimrit Shtern , Mathias Staudigl

This paper is concerned with the large-scale regularity in the homogenization of elliptic systems of elasticity with periodic high-contrast coefficients. We obtain the large-scale Lipschitz estimate that is uniform with respect to the…

Analysis of PDEs · Mathematics 2020-08-12 Zhongwei Shen

Adjusting the learning rate schedule in stochastic gradient methods is an important unresolved problem which requires tuning in practice. If certain parameters of the loss function such as smoothness or strong convexity constants are known,…

Machine Learning · Statistics 2020-11-23 Xiaoxia Wu , Rachel Ward , Léon Bottou

We introduce a new concept, data irrecoverability, and show that the well-studied concept of data privacy is sufficient but not necessary for data irrecoverability. We show that there are several regularized loss minimization problems that…

Machine Learning · Computer Science 2021-07-07 Zitao Li , Jean Honorio

We develop a higher regularity theory for general quasilinear elliptic equations and systems in divergence form with random coefficients. The main result is a large-scale $L^\infty$-type estimate for the gradient of a solution. The estimate…

Analysis of PDEs · Mathematics 2016-01-27 Scott N. Armstrong , Jean-Christophe Mourrat
‹ Prev 1 4 5 6 7 8 10 Next ›