中文
相关论文

相关论文: Exactly Computing the Local Lipschitz Constant of …

200 篇论文

We initiate a program of average smoothness analysis for efficiently learning real-valued functions on metric spaces. Rather than using the Lipschitz constant as the regularizer, we define a local slope at each point and gauge the function…

统计理论 · 数学 2020-11-10 Yair Ashlagi , Lee-Ad Gottlieb , Aryeh Kontorovich

Attention is a powerful component of modern neural networks across a wide variety of domains. In this paper, we seek to quantify the regularity (i.e. the amount of smoothness) of the attention operation. To accomplish this goal, we propose…

机器学习 · 统计学 2021-02-11 James Vuckovic , Aristide Baratin , Remi Tachet des Combes

We introduce the Lipschitz matrix: a generalization of the scalar Lipschitz constant for functions with many inputs. Among the Lipschitz matrices compatible a particular function, we choose the smallest such matrix in the Frobenius norm to…

数值分析 · 数学 2023-03-24 Jeffrey M. Hokanson , Paul G. Constantine

The goal of the paper is to design sequential strategies which lead to efficient optimization of an unknown function under the only assumption that it has a finite Lipschitz constant. We first identify sufficient conditions for the…

机器学习 · 统计学 2017-06-19 Cédric Malherbe , Nicolas Vayatis

Recent research has established that the local Lipschitz constant of a neural network directly influences its adversarial robustness. We exploit this relationship to construct an ensemble of neural networks which not only improves the…

机器学习 · 计算机科学 2022-06-02 Thulasi Tholeti , Sheetal Kalyani

The convergence theory for the gradient sampling algorithm is extended to directionally Lipschitz functions. Although directionally Lipschitz functions are not necessarily locally Lipschitz, they are almost everywhere differentiable and…

最优化与控制 · 数学 2021-07-13 James V. Burke , Qiuying Lin

In this paper, we consider the inverse problem of detecting a corrosion coefficient between two layers of a conducting medium from the Neumann-to-Dirichlet map. This inverse problem is motivated by the description of the index of corrosion…

数值分析 · 数学 2019-04-08 Bastian Harrach , Houcine Meftahi

We consider the problem of training a deep neural network with nonsmooth regularization to retrieve a sparse and efficient sub-structure. Our regularizer is only assumed to be lower semi-continuous and prox-bounded. We combine an adaptive…

机器学习 · 统计学 2022-06-20 Dounia Lakhmiri , Dominique Orban , Andrea Lodi

In many real-world applications, it is undesirable to drastically change the problem solution after a small perturbation in the input, as unstable outputs can lead to costly transaction fees, privacy and security concerns, reduced user…

数据结构与算法 · 计算机科学 2025-11-11 Quanquan C. Liu , Grigoris Velegkas , Yuichi Yoshida , Felix Zhou

We consider a stochastic version of the proximal point algorithm for optimization problems posed on a Hilbert space. A typical application of this is supervised learning. While the method is not new, it has not been extensively analyzed in…

最优化与控制 · 数学 2021-09-28 Monika Eisenmann , Tony Stillfjord , Måns Williamson

Using uniform global Carleman estimates for discrete elliptic and semi-discrete hyperbolic equations, we study Lipschitz and logarithmic stability for the inverse problem of recovering a potential in a semi-discrete wave equation,…

偏微分方程分析 · 数学 2014-09-29 Lucie Baudouin , Sylvain Ervedoza , Axel Osses

We study the type of solutions to which stochastic gradient descent converges when used to train a single hidden-layer multivariate ReLU network with the quadratic loss. Our results are based on a dynamical stability analysis. In the…

机器学习 · 计算机科学 2023-07-03 Mor Shpigel Nacson , Rotem Mulayoff , Greg Ongie , Tomer Michaeli , Daniel Soudry

In this paper we are concerned with the approximation of functions by single hidden layer neural networks with ReLU activation functions on the unit circle. In particular, we are interested in the case when the number of data-points exceeds…

偏微分方程分析 · 数学 2021-04-02 Benny Avelin , Vesa Julin

The point estimates of ReLU classification networks---arguably the most widely used neural network architecture---have been shown to yield arbitrarily high confidence far away from the training data. This architecture, in conjunction with a…

机器学习 · 统计学 2020-07-20 Agustinus Kristiadi , Matthias Hein , Philipp Hennig

We present PSiLON Net, an MLP architecture that uses $L_1$ weight normalization for each weight vector and shares the length parameter across the layer. The 1-path-norm provides a bound for the Lipschitz constant of a neural network and…

机器学习 · 计算机科学 2024-05-01 Aditya Biswas

Deep neural networks have shown remarkable performance across a wide range of vision-based tasks, particularly due to the availability of large-scale datasets for training and better architectures. However, data seen in the real world are…

机器学习 · 计算机科学 2018-11-26 Muhammad Usama , Dong Eui Chang

In this paper, we propose an inexact proximal Newton-type method for nonconvex composite problems. We establish the global convergence rate of the order $\mathcal{O}(k^{-1/2})$ in terms of the minimal norm of the KKT residual mapping and…

最优化与控制 · 数学 2024-12-26 Hong Zhu

We establish Lipschitz stability properties for a class of inverse problems. In that class, the associated direct problem is formulated by an integral operator Am depending non-linearly on a parameter m and operating on a function u. In the…

数值分析 · 数学 2023-02-27 Darko Volkov

In this paper, we consider a class of structured nonconvex nonsmooth optimization problems, in which the objective function is formed by the sum of a possibly nonsmooth nonconvex function and a differentiable function whose gradient is…

最优化与控制 · 数学 2024-10-01 Tan Nhat Pham , Minh N. Dao , Rakibuzzaman Shah , Nargiz Sultanova , Guoyin Li , Syed Islam

We study unconstrained Online Linear Optimization with Lipschitz losses. Motivated by the pursuit of instance optimality, we propose a new algorithm that simultaneously achieves ($i$) the AdaGrad-style second order gradient adaptivity; and…

机器学习 · 计算机科学 2024-02-23 Zhiyu Zhang , Heng Yang , Ashok Cutkosky , Ioannis Ch. Paschalidis