English
Related papers

Related papers: Deep JKO: time-implicit particle methods for gener…

200 papers

We study the convergence of gradient flow for the training of deep neural networks. If Residual Neural Networks are a popular example of very deep architectures, their training constitutes a challenging optimization problem due notably to…

Machine Learning · Computer Science 2025-07-22 Raphaël Barboni , Gabriel Peyré , François-Xavier Vialard

The theory of Wasserstein gradient flows in the space of probability measures has made an enormous progress over the last twenty years. It constitutes a unified and powerful framework in the study of dissipative partial differential…

Analysis of PDEs · Mathematics 2022-01-17 Daniel Adams , Manh Hong Duong , Goncalo dos Reis

In this paper, we introduce a modular deep neural network (DNN) framework for data-driven reduced order modeling of dynamical systems relevant to fluid flows. We propose various deep neural network architectures which numerically predict…

Computational Physics · Physics 2019-09-04 S. Pawar , S. M. Rahman , H. Vaddireddy , O. San , A. Rasheed , P. Vedula

In this paper we bring together some of the key ideas and methods of two disparate fields of mathematical research, frame theory and optimal transport, using the methods of the second to answer questions posed in the first. In particular,…

Functional Analysis · Mathematics 2022-12-01 Clare Wickman , Kasso Okoudjou

This paper provides a new avenue for exploiting deep neural networks to improve physics-based simulation. Specifically, we integrate the classic Lagrangian mechanics with a deep autoencoder to accelerate elastic simulation of deformable…

Machine Learning · Computer Science 2021-02-23 Siyuan Shen , Yang Yin , Tianjia Shao , He Wang , Chenfanfu Jiang , Lei Lan , Kun Zhou

We prove convergence of a variational formulation of the BDF2 method applied to the non-linear Fokker-Planck equation. Our approach is inspired by the JKO-method and exploits the differential structure of the underlying $L^2$-Wasserstein…

Numerical Analysis · Mathematics 2018-01-30 Simon Plazotta

We establish the gradient flow representation of diffusion with mobility $b$ with respect to the modified Wasserstein quasi-metric $W_h$, where $h(r)=rb(r)$. The appropriate selection of the free energy functional depends on the specific…

Probability · Mathematics 2025-01-22 Zhenxin Liu , Xuewei Wang

In this paper we introduce a theoretical framework for semi-discrete optimization using ideas from optimal transport. Our primary motivation is in the field of deep learning, and specifically in the task of neural architecture search. With…

Analysis of PDEs · Mathematics 2022-02-01 Nicolas Garcia Trillos , Javier Morales

This paper presents a Wasserstein attraction approach for solving dynamic mass transport problems over networks. In the transport problem over networks, we start with a distribution over the set of nodes that needs to be "transported" to a…

Optimization and Control · Mathematics 2022-04-28 Ferran Arqué , César A. Uribe , Carlos Ocampo-Martinez

This study focuses on a Wasserstein-type gradient flow, which represents an optimization process of a continuous model of a Deep Neural Network (DNN). First, we establish the existence of a minimizer for an average loss of the model under…

Machine Learning · Computer Science 2024-04-16 Noboru Isobe

Optimal experimental design (OED) aims to choose the observations in an experiment to be as informative as possible, according to certain statistical criteria. In the linear case (when the observations depend linearly on the unknown…

Numerical Analysis · Mathematics 2026-02-25 Ruhui Jin , Martin Guerra , Qin Li , Stephen Wright

In this work, we propose a numerical method to compute the Wasserstein Hamiltonian flow (WHF), which is a Hamiltonian system on the probability density manifold. Many well-known PDE systems can be reformulated as WHFs. We use parameterized…

Numerical Analysis · Mathematics 2023-07-18 Hao Wu , Shu Liu , Xiaojing Ye , Haomin Zhou

We propose efficient numerical schemes for implementing the natural gradient descent (NGD) for a broad range of metric spaces with applications to PDE-based optimization problems. Our technique represents the natural gradient direction as a…

Optimization and Control · Mathematics 2023-01-12 Levon Nurbekyan , Wanzhou Lei , Yunan Yang

We consider a minimax problem motivated by distributionally robust optimization (DRO) when the worst-case distribution is continuous, leading to significant computational challenges due to the infinite-dimensional nature of the optimization…

Machine Learning · Statistics 2024-12-31 Linglingzhi Zhu , Yao Xie

Variational inference (VI) can be cast as an optimization problem in which the variational parameters are tuned to closely align a variational distribution with the true posterior. The optimization task can be approached through vanilla…

Machine Learning · Computer Science 2025-04-24 Dai Hai Nguyen , Tetsuya Sakurai , Hiroshi Mamitsuka

Variational inference is a technique that approximates a target distribution by optimizing within the parameter space of variational families. On the other hand, Wasserstein gradient flows describe optimization within the space of…

Machine Learning · Statistics 2023-11-01 Mingxuan Yi , Song Liu

Model reduction for fluid flow simulation continues to be of great interest across a number of scientific and engineering fields. Here, we explore the use of Neural Ordinary Differential Equations, a recently introduced family of…

Machine Learning · Computer Science 2021-04-30 Sourav Dutta , Peter Rivera-Casillas , Matthew W. Farthing

Solving Fredholm equations of the first kind is crucial in many areas of the applied sciences. In this work we adopt a probabilistic and variational point of view by considering a minimization problem in the space of probability measures…

Optimization and Control · Mathematics 2024-05-17 Francesca R. Crucinio , Valentin De Bortoli , Arnaud Doucet , Adam M. Johansen

This is an expository paper on the theory of gradient flows, and in particular of those PDEs which can be interpreted as gradient flows for the Wasserstein metric on the space of probability measures (a distance induced by optimal…

Analysis of PDEs · Mathematics 2016-09-14 Filippo Santambrogio

We study the implicit bias of gradient flow (i.e., gradient descent with infinitesimal step size) on linear neural network training. We propose a tensor formulation of neural networks that includes fully-connected, diagonal, and…

Machine Learning · Computer Science 2021-09-13 Chulhee Yun , Shankar Krishnan , Hossein Mobahi
‹ Prev 1 3 4 5 6 7 10 Next ›