English
Related papers

Related papers: Atomic Gradient Flows: Gradient Flows on Sparse Re…

200 papers

We study the quantitative convergence of drift-diffusion PDEs that arise as Wasserstein gradient flows of linearly convex functions over the space of probability measures on ${\mathbb R}^d$. In this setting, the objective is in general not…

Optimization and Control · Mathematics 2025-07-17 Lénaïc Chizat , Maria Colombo , Xavier Fernández-Real

We study the equation of one-dimensional quasistatic nonlinear viscoelasticity with Dirichlet boundary conditions, in the particular case that the underlying dissipation geometry (provided by the viscosity) is comparable to the Bhattacharya…

Analysis of PDEs · Mathematics 2026-05-12 Alexander Mielke , Billy Sumners

The theory of Wasserstein gradient flows in the space of probability measures has made an enormous progress over the last twenty years. It constitutes a unified and powerful framework in the study of dissipative partial differential…

Analysis of PDEs · Mathematics 2022-01-17 Daniel Adams , Manh Hong Duong , Goncalo dos Reis

What features neural networks learn, and how, remains an open question. In this paper, we introduce Alternating Gradient Flows (AGF), an algorithmic framework that describes the dynamics of feature learning in two-layer networks trained…

The $L^2$ gradient flow of the Ginzburg-Landau free energy functional leads to the Allen Cahn equation that is widely used for modeling phase separation. Machine learning methods for solving the Allen-Cahn equation in its strong form suffer…

Machine Learning · Computer Science 2025-03-27 Revanth Mattey , Susanta Ghosh

Training sparse networks to converge to the same performance as dense neural architectures has proven to be elusive. Recent work suggests that initialization is the key. However, while this direction of research has had some success,…

Machine Learning · Computer Science 2021-06-17 Kale-ab Tessera , Sara Hooker , Benjamin Rosman

Minimizing functionals in the space of probability distributions can be done with Wasserstein gradient flows. To solve them numerically, a possible approach is to rely on the Jordan-Kinderlehrer-Otto (JKO) scheme which is analogous to the…

Machine Learning · Computer Science 2022-11-16 Clément Bonet , Nicolas Courty , François Septier , Lucas Drumetz

We study optimization problems whereby the optimization variable is a probability measure. Since the probability space is not a vector space, many classical and powerful methods for optimization (e.g., gradients) are of little help. Thus,…

Optimization and Control · Mathematics 2024-06-18 Nicolas Lanzetti , Antonio Terpin , Florian Dörfler

Wasserstein distributionally robust optimization offers a framework for model fitting in machine learning under potential shifts in the data distribution. We study a regularized variant of this problem in which entropic smoothing produces a…

Optimization and Control · Mathematics 2026-05-28 Tam Le

By building upon the recent theory that established the connection between implicit generative modeling (IGM) and optimal transport, in this study, we propose a novel parameter-free algorithm for learning the underlying distributions of…

Machine Learning · Statistics 2019-06-12 Antoine Liutkus , Umut Şimşekli , Szymon Majewski , Alain Durmus , Fabian-Robert Stöter

Nonsmooth nonconvex optimization problems broadly emerge in machine learning and business decision making, whereas two core challenges impede the development of efficient solution methods with finite-time convergence guarantee: the lack of…

Optimization and Control · Mathematics 2022-10-18 Tianyi Lin , Zeyu Zheng , Michael I. Jordan

Stochastic optimization plays a crucial role in the advancement of deep learning technologies. Over the decades, significant effort has been dedicated to improving the training efficiency and robustness of deep neural networks, via various…

Machine Learning · Computer Science 2024-08-21 Huixiu Jiang , Ling Yang , Yu Bao , Rutong Si , Sikun Yang

Flow matching (FM) learns vector fields by regressing stochastic velocity targets along intermediate distributions $p_t$. We identify a geometric optimization bottleneck in this regression problem: when the covariance $\Sigma_t$ of $p_t$ is…

Machine Learning · Computer Science 2026-05-14 Shadab Ahamed , Eshed Gal , Md Shahriar Rahim Siddiqui , Simon Ghyselincks , Moshe Eliasof , Eldad Haber

We design and compute first-order implicit-in-time variational schemes with high-order spatial discretization for initial value gradient flows in generalized optimal transport metric spaces. We first review some examples of gradient flows…

Numerical Analysis · Mathematics 2023-08-16 Guosheng Fu , Stanley Osher , Wuchen Li

This paper is a survey of the generalized Hamiltonian gradient flow (GHGF) framework for Hamilton-Jacobi equations, with an emphasis on the propagation of singularities and its connections to weak KAM theory, optimal transport and mean…

Analysis of PDEs · Mathematics 2026-05-07 Wei Cheng , Jiahui Hong

In this paper, a projected primal-dual gradient flow of augmented Lagrangian is presented to solve convex optimization problems that are not necessarily strictly convex. The optimization variables are restricted by a convex set with…

Optimization and Control · Mathematics 2018-10-31 Han Zhang , Jieqiang Wei , Peng Yi , Xiaoming Hu

Vision-Language Latent Diffusion Models (LDMs) (Rombach et al., 2022) provide powerful generative priors for inverse problems. However, existing LDM-based inverse solvers typically require a large number of neural function evaluations…

Machine Learning · Statistics 2026-05-11 Alessio Spagnoletti , Tim Y. J. Wang , Marcelo Pereyra , O. Deniz Akyildiz

We introduce adaptive, tuning-free step size schedules for gradient-based sampling algorithms obtained as time-discretizations of Wasserstein gradient flows. The result is a suite of tuning-free sampling algorithms, including tuning-free…

Methodology · Statistics 2025-10-30 Louis Sharrock , Christopher Nemeth

This paper presents a groundbreaking approach to causal inference by integrating continuous normalizing flows (CNFs) with parametric submodels, enhancing their geometric sensitivity and improving upon traditional Targeted Maximum Likelihood…

Machine Learning · Computer Science 2024-02-02 Kaiwen Hou

In this work we propose a differential geometric motivation for Nesterov's accelerated gradient method (AGM) for strongly-convex problems. By considering the optimization procedure as occurring on a Riemannian manifold with a natural…

Machine Learning · Computer Science 2019-11-21 Aaron Defazio
‹ Prev 1 3 4 5 6 7 10 Next ›