English
Related papers

Related papers: On the error estimate of gradient inclusions

200 papers

We derive error estimates for the piecewise linear finite element approximation of the Laplace--Beltrami operator on a bounded, orientable, $C^3$, surface without boundary on general shape regular meshes. As an application, we consider a…

Numerical Analysis · Mathematics 2017-05-15 Johnny Guzman , Alexandre Madureira , Marcus Sarkis , Shawn Walker

We study the convergence issue for the gradient algorithm (employing general step sizes) for optimization problems on general Riemannian manifolds (without curvature constraints). Under the assumption of the local convexity/quasi-convexity…

Optimization and Control · Mathematics 2019-10-08 Chong Li , Xiangmei Wang , Jinhua Wang , Jen-Chih Yao

A non-negative integer invariant, estimating from below the number of geometrically different critical points of a smooth function $f$ defined in the 2-disk, $f:\mathbb{B}^{2}\rightarrow\mathbb{R}$, is considered. (We denote it by…

Geometric Topology · Mathematics 2018-10-10 Simeon Stefanov

Many statistical $M$-estimators are based on convex optimization problems formed by the combination of a data-dependent loss function with a norm-based regularizer. We analyze the convergence rates of projected gradient and composite…

Machine Learning · Statistics 2012-07-26 Alekh Agarwal , Sahand N. Negahban , Martin J. Wainwright

This paper examines the problem of locating outlier columns in a large, otherwise low-rank, matrix. We propose a simple two-step adaptive sensing and inference approach and establish theoretical guarantees for its performance; our results…

Information Theory · Computer Science 2015-06-22 Xingguo Li , Jarvis Haupt

The demands of accuracy in measurements and engineering models today, renders the condition number of problems larger. While a corresponding increase in the precision of floating point numbers ensured a stable computing, the uncertainty in…

Numerical Analysis · Mathematics 2022-09-12 Puneet Jain , Krishna Manglani , Murugesan Venkatapathi

In this paper, we analyze the accuracy of gradient estimates obtained by linear interpolation when the underlying function is subject to bounded measurement noise. The total gradient error is decomposed into a deterministic component…

Numerical Analysis · Mathematics 2025-07-29 Alejandro G. Marchetti , Dominique Bonvin

Generalized linear models (GLMs) arise in high-dimensional machine learning, statistics, communications and signal processing. In this paper we analyze GLMs when the data matrix is random, as relevant in problems such as compressed sensing,…

Information Theory · Computer Science 2019-04-01 Jean Barbier , Florent Krzakala , Nicolas Macris , Léo Miolane , Lenka Zdeborová

Autoencoders are a popular model in many branches of machine learning and lossy data compression. However, their fundamental limits, the performance of gradient methods and the features learnt during optimization remain poorly understood,…

Machine Learning · Computer Science 2023-01-12 Alexander Shevchenko , Kevin Kögler , Hamed Hassani , Marco Mondelli

In this paper, we provide the universal first-order methods of Composite Optimization with new complexity analysis. It delivers some universal convergence guarantees, which are not linked directly to any parametric problem class. However,…

Optimization and Control · Mathematics 2025-09-26 Yurii Nesterov

We review the literature on algorithms for estimating the index space in a multi-index model. The primary focus is on computationally efficient (polynomial-time) algorithms in Gaussian space, the assumptions under which consistency is…

Machine Learning · Statistics 2025-06-17 Joan Bruna , Daniel Hsu

An ensemble method is introduced that utilizes randomization and loss function gradients to compute a prediction. Multiple weakly-correlated estimators approximate the gradient at randomly sampled points on the error surface and are…

Machine Learning · Computer Science 2020-09-15 Nicholas Smith

A formula is given for the propagation of errors during matrix inversion. An explicit calculation for a 2 by 2 matrix using both the formula and a Monte Carlo calculation are compared. A prescription is given to determine when a matrix with…

High Energy Physics - Experiment · Physics 2009-10-31 M. Lefebvre , R. K. Keeler , R. Sobie , J. White

The optimization algorithms are crucial in training physics-informed neural networks (PINNs), as unsuitable methods may lead to poor solutions. Compared to the common gradient descent (GD) algorithm, implicit gradient descent (IGD)…

Machine Learning · Computer Science 2025-08-04 Xianliang Xu , Ting Du , Wang Kong , Bin Shan , Ye Li , Zhongyi Huang

For real symmetric matrices that are accessible only through matrix vector products, we present Monte Carlo estimators for computing the diagonal elements. Our probabilistic bounds for normwise absolute and relative errors apply to Monte…

Numerical Analysis · Mathematics 2022-03-18 Eric Hallman , Ilse C. F. Ipsen , Arvind Saibaba

In this paper, we present several estimators of the diagonal elements of the inverse of the covariance matrix, called precision matrix, of a sample of iid random vectors. The focus is on high dimensional vectors having a sparse precision…

Statistics Theory · Mathematics 2017-07-31 Samuel Balmand , Arnak S. Dalalyan

The framework of differential inclusions encompasses modern optimal control and the calculus of variations. Necessary optimality conditions in the literature identify potentially optimal paths, but do not show how to perturb paths to…

Optimization and Control · Mathematics 2012-05-01 C. H. Jeffrey Pang

We study a version of the proximal gradient algorithm for which the gradient is intractable and is approximated by Monte Carlo methods (and in particular Markov Chain Monte Carlo). We derive conditions on the step size and the Monte Carlo…

Statistics Theory · Mathematics 2016-11-22 Yves F. Atchade , Gersende Fort , Eric Moulines

Partitioning a set of elements into an unknown number of mutually exclusive subsets is essential in many machine learning problems. However, assigning elements, such as samples in a dataset or neurons in a network layer, to an unknown and…

Machine Learning · Computer Science 2023-11-10 Thomas M. Sutter , Alain Ryser , Joram Liebeskind , Julia E. Vogt

This paper studies global a priori gradient estimates for divergence-type equations patterned over the $p$-Laplacian with first-order terms having polynomial growth with respect to the gradient, under suitable integrability assumptions on…

Analysis of PDEs · Mathematics 2024-10-22 Marco Cirant , Alessandro Goffi , Tommaso Leonori