English
Related papers

Related papers: Normalized Gradients for All

200 papers

In this paper, we propose a generalized conditional gradient method for multiobjective optimization, which can be viewed as an improved extension of the classical Frank-Wolfe (conditional gradient) method for single-objective optimization.…

Optimization and Control · Mathematics 2025-03-25 Anteneh Getachew Gebrie , Ellen Hidemi Fukuda

Randomized smoothing is a widely adopted technique for optimizing nonsmooth objective functions. However, its efficiency analysis typically relies on global Lipschitz continuity, a condition rarely met in practical applications. To address…

Optimization and Control · Mathematics 2025-09-10 Jingfan Xia , Zhenwei Lin , Qi Deng

$L_0$-smoothness, which has been pivotal to advancing decentralized optimization theory, is often fairly restrictive for modern tasks like deep learning. The recent advent of relaxed $(L_0,L_1)$-smoothness condition enables improved…

Optimization and Control · Mathematics 2025-08-13 Zhanhong Jiang , Aditya Balu , Soumik Sarkar

Machine learning methods have seen a meteoric rise in their applications in the scientific community. However, little effort has been put into understanding these "black box" models. We show how one can apply integrated gradients (IGs) to…

Machine Learning · Computer Science 2024-12-19 Jai Bardhan , Cyrin Neeraj , Mihir Rawat , Subhadip Mitra

In this paper we study on smooth bounded domains the global regularity (up to the boundary) for weak solutions to systems having $p$-structure depending only on the symmetric part of the gradient.

Analysis of PDEs · Mathematics 2019-05-13 Luigi C. Berselli , Michael Ruzicka

The (gradient-based) bilevel programming framework is widely used in hyperparameter optimization and has achieved excellent performance empirically. Previous theoretical work mainly focuses on its optimization properties, while leaving the…

Machine Learning · Computer Science 2021-10-26 Fan Bao , Guoqiang Wu , Chongxuan Li , Jun Zhu , Bo Zhang

Convolutions encode equivariance symmetries into neural networks leading to better generalisation performance. However, symmetries provide fixed hard constraints on the functions a network can represent, need to be specified in advance, and…

Machine Learning · Computer Science 2023-10-11 Tycho F. A. van der Ouderaa , Alexander Immer , Mark van der Wilk

The norm in classical Sobolev spaces can be expressed as a difference quotient. This expression can be used to generalize the space to the fractional smoothness case. Because the difference quotient is based on shifting the function, it…

Functional Analysis · Mathematics 2020-08-06 Rita Ferreira , Peter Hästö , Ana Margarida Ribeiro

In this paper, acceleration of gradient methods for convex optimization problems with weak levels of convexity and smoothness is considered. Starting from the universal fast gradient method which was designed to be an optimal method for…

Optimization and Control · Mathematics 2022-06-10 Jongho Park

In this thesis we develop a novel framework to study smooth and strongly convex optimization algorithms, both deterministic and stochastic. Focusing on quadratic functions we are able to examine optimization algorithms as a recursive…

Optimization and Control · Mathematics 2014-10-24 Yossi Arjevani

This paper considers optimization of smooth nonconvex functionals in smooth infinite dimensional spaces. A H\"older gradient descent algorithm is first proposed for finding approximate first-order points of regularized polynomial…

Optimization and Control · Mathematics 2021-04-07 Serge Gratton , Sadok Jerad , Philippe L. Toint

Normalizing flows are a powerful technique for obtaining reparameterizable samples from complex multimodal distributions. Unfortunately current approaches fall short when the underlying space has a non trivial topology, and are only…

Machine Learning · Statistics 2020-06-12 Luca Falorsi , Patrick Forré

We develop a unified approach to universality of local scaling limits for eigenvalues of random normal matrices, or equivalently for planar Coulomb gases at inverse temperature $\beta=2$. The approach is direct in that it does not rely on…

Probability · Mathematics 2025-11-25 Joakim Cronvall , Aron Wennman

Due to its applications in many different places in machine learning and other connected engineering applications, the problem of minimization of a smooth function that satisfies the Polyak-{\L}ojasiewicz condition receives much attention…

Optimization and Control · Mathematics 2022-12-09 Ilya A. Kuruzov , Fedor S. Stonyakin , Mohammad S. Alkousa

Gradient clipping has long been considered essential for ensuring the convergence of Stochastic Gradient Descent (SGD) in the presence of heavy-tailed gradient noise. In this paper, we revisit this belief and explore whether gradient…

Machine Learning · Computer Science 2025-11-20 Tao Sun , Xinwang Liu , Kun Yuan

This paper is devoted to the study of the solution of a stochastic convex black box optimization problem. Where the black box problem means that the gradient-free oracle only returns the value of objective function, not its gradient. We…

Optimization and Control · Mathematics 2023-04-18 Aleksandr Lobanov

Although it is relatively easy to apply, the gradient method often displays a disappointingly slow rate of convergence. Its convergence is specially based on the structure of the matrix of the algebraic linear system, and on the choice of…

Numerical Analysis · Mathematics 2025-06-03 Ibrahima Dione

In this paper, via applying the method developed by A. Cianchi and V. Maz'ya, the author obtains the global boundedness of the gradient for solutions to Dirichlet and Neumann problems of a class of Schr\"odinger equations under the minimal…

Analysis of PDEs · Mathematics 2016-03-01 Sibei Yang

We study smoothness of generalized solutions of nonlocal elliptic problems in plane bounded domains with piecewise smooth boundary. The case where the support of nonlocal terms can intersect the boundary is considered. We announce…

Analysis of PDEs · Mathematics 2014-04-22 Pavel Gurevich

Adaptive gradient methods like AdaGrad are widely used in optimizing neural networks. Yet, existing convergence guarantees for adaptive gradient methods require either convexity or smoothness, and, in the smooth setting, only guarantee…

Machine Learning · Computer Science 2019-10-22 Xiaoxia Wu , Simon S. Du , Rachel Ward