English
Related papers

Related papers: Deep JKO: time-implicit particle methods for gener…

200 papers

We propose a fully discrete variational scheme for nonlinear evolution equations with gradient flow structure on the space of finite Radon measures on an interval with respect to a generalized version of the Wasserstein distance with…

Numerical Analysis · Mathematics 2016-09-29 Jonathan Zinsl , Daniel Matthes

We present a structure-preserving Eulerian algorithm for solving $L^2$-gradient flows and a structure-preserving Lagrangian algorithm for solving generalized diffusions. Both algorithms employ neural networks as tools for spatial…

Numerical Analysis · Mathematics 2024-04-16 Ziqing Hu , Chun Liu , Yiwei Wang , Zhiliang Xu

Even for the gradient descent (GD) method applied to neural network training, understanding its optimization dynamics, including convergence rate, iterate trajectories, function value oscillations, and especially its implicit acceleration,…

Machine Learning · Computer Science 2026-05-22 Alexander Tyurin

We introduce the so called DeepParticle method to learn and generate invariant measures of stochastic dynamical systems with physical parameters based on data computed from an interacting particle method (IPM). We utilize the expressiveness…

Machine Learning · Computer Science 2022-06-22 Zhongjian Wang , Jack Xin , Zhiwen Zhang

We study fully discrete linearized Galerkin finite element approximations to a nonlinear gradient flow, applications of which can be found in many areas. Due to the strong nonlinearity of the equation, existing analyses for implicit schemes…

Numerical Analysis · Mathematics 2014-06-17 Buyang Li , Weiwei Sun

Applications such as adversarially robust training and Wasserstein Distributionally Robust Optimization (WDRO) can be naturally formulated as min-sum-max optimization problems. While this formulation can be rewritten as an equivalent…

Optimization and Control · Mathematics 2025-02-26 Wei Liu , Muhammad Khan , Gabriel Mancino-Ball , Yangyang Xu

Recent results have shown that for two-layer fully connected neural networks, gradient flow converges to a global optimum in the infinite width limit, by making a connection between the mean field dynamics and the Wasserstein gradient flow.…

Optimization and Control · Mathematics 2020-07-16 Walid Krichene , Kenneth F. Caluya , Abhishek Halder

In this paper, we study the stochastic Hamiltonian flow in Wasserstein manifold, the probability density space equipped with $L^2$-Wasserstein metric tensor, via the Wong--Zakai approximation. We begin our investigation by showing that the…

Probability · Mathematics 2021-12-01 Jianbo Cui , Shu Liu , Haomin Zhou

Recent works on optical flow estimation use neural networks to predict the flow field that maps positions of one image to positions of the other. These networks consist of a feature extractor, a correlation volume, and finally several…

Computer Vision and Pattern Recognition · Computer Science 2025-06-05 Leyla Mirvakhabova , Hong Cai , Jisoo Jeong , Hanno Ackermann , Farhad Zanjani , Fatih Porikli

In view of training increasingly complex learning architectures, we establish a nonsmooth implicit function theorem with an operational calculus. Our result applies to most practical problems (i.e., definable problems) provided that a…

Machine Learning · Computer Science 2022-04-06 Jérôme Bolte , Tam Le , Edouard Pauwels , Antonio Silveti-Falls

This article provides a comprehensive understanding of optimization in deep learning, with a primary focus on the challenges of gradient vanishing and gradient exploding, which normally lead to diminished model representational ability and…

Machine Learning · Computer Science 2023-11-14 Xianbiao Qi , Jianan Wang , Lei Zhang

Differential equations in general and neural ODEs in particular are an essential technique in continuous-time system identification. While many deterministic learning algorithms have been designed based on numerical integration via the…

Machine Learning · Computer Science 2021-10-18 Lenart Treven , Philippe Wenk , Florian Dörfler , Andreas Krause

It is well-known that many diffusion equations can be recast as Wasserstein gradient flows. Moreover, in recent years, by modifying the Wasserstein distance appropriately, this technique has been transferred to further evolution equations…

Probability · Mathematics 2020-10-15 Kaveh Bashiri , Anton Bovier

This paper presents a new gradient flow dissipation geometry over non-negative and probability measures. This is motivated by a principled construction that combines the unbalanced optimal transport and interaction forces modeled by…

Machine Learning · Computer Science 2024-11-01 Egor Gladin , Pavel Dvurechensky , Alexander Mielke , Jia-Jie Zhu

We introduce a new family of deep neural network models. Instead of specifying a discrete sequence of hidden layers, we parameterize the derivative of the hidden state using a neural network. The output of the network is computed using a…

Machine Learning · Computer Science 2019-12-17 Ricky T. Q. Chen , Yulia Rubanova , Jesse Bettencourt , David Duvenaud

We present practical Levenberg-Marquardt variants of Gauss-Newton and natural gradient methods for solving non-convex optimization problems that arise in training deep neural networks involving enormous numbers of variables and huge data…

Machine Learning · Computer Science 2019-06-07 Yi Ren , Donald Goldfarb

Latent-variable energy-based models (LVEBMs) assign a single normalized energy to joint pairs of observed data and latent variables, offering expressive generative modeling while capturing hidden structure. We recast maximum-likelihood…

Machine Learning · Computer Science 2025-10-20 Shiqin Tang , Shuxin Zhuang , Rong Feng , Runsheng Yu , Hongzong Li , Youzhi Zhang

We consider a class of optimization problems on the space of probability measures motivated by the mean-field approach to studying neural networks. Such problems can be solved by constructing continuous-time gradient flows that converge to…

Optimization and Control · Mathematics 2026-02-18 Petra Lazić , Linshan Liu , Mateusz B. Majka

Many machine learning problems can be seen as approximating a \textit{target} distribution using a \textit{particle} distribution by minimizing their statistical discrepancy. Wasserstein Gradient Flow can move particles along a path that…

Machine Learning · Statistics 2024-06-07 Song Liu , Jiahao Yu , Jack Simons , Mingxuan Yi , Mark Beaumont

We examine the infinite-dimensional optimization problem of finding a decomposition of a probability measure into K probability sub-measures to minimize specific loss functions inspired by applications in clustering and user grouping. We…

Optimization and Control · Mathematics 2024-06-04 Jiangze Han , Christopher Thomas Ryan , Xin T. Tong