English
Related papers

Related papers: The Hamilton-Jacobi Theory of Deep Learning

200 papers

We design fast numerical methods for Hamilton-Jacobi equations in density space (HJD), which arises in optimal transport and mean field games. We overcome the curse-of-infinite-dimensionality nature of HJD by proposing a generalized Hopf…

Numerical Analysis · Mathematics 2018-05-07 Yat Tin Chow , Wuchen Li , Stanley Osher , Wotao Yin

We study the connection between the highly non-convex loss function of a simple model of the fully-connected feed-forward neural network and the Hamiltonian of the spherical spin-glass model under the assumptions of: i) variable…

Machine Learning · Computer Science 2015-01-23 Anna Choromanska , Mikael Henaff , Michael Mathieu , Gérard Ben Arous , Yann LeCun

This paper studies Hamilton-Jacobi equations of evolution type defined in a general metric space. We give a notion of a solution through optimal principles and establish a unique existence theorem of the solution for initial value problems.…

Analysis of PDEs · Mathematics 2014-07-30 Atsushi Nakayasu

This note lays part of the theoretical ground for a definition of differential systems modeling reinforcement learning in continuous time non-Markovian rough environments. Specifically we focus on optimal relaxed control of rough equations…

Optimization and Control · Mathematics 2024-02-29 Prakash Chakraborty , Harsha Honnappa , Samy Tindel

Finding the optimal hyperparameters of a model can be cast as a bilevel optimization problem, typically solved using zero-order techniques. In this work we study first-order methods when the inner optimization problem is convex but…

We study a selection problem for degenerate viscous Hamilton--Jacobi equations with convex Hamiltonians, in which the approximation procedure combines a nonlinear discounted approximation with a small potential perturbation. A key question…

Analysis of PDEs · Mathematics 2026-05-14 Qinbo Chen , Zhi-Xiang Zhu

Gradient descent during the learning process of a neural network can be subject to many instabilities. The spectral density of the Jacobian is a key component for analyzing stability. Following the works of Pennington et al., such Jacobians…

Machine Learning · Statistics 2023-04-26 Reda Chhaibi , Tariq Daouda , Ezechiel Kahn

We argue that Hamilton-Jacobi equations provide a convenient and intuitive approach for studying the large-scale behavior of mean-field disordered systems. This point of view is illustrated on the problem of inference of a rank-one matrix.…

Probability · Mathematics 2018-11-13 Jean-Christophe Mourrat

Many imaging problems can be formulated as inverse problems expressed as finite-dimensional optimization problems. These optimization problems generally consist of minimizing the sum of a data fidelity and regularization terms. In [23,26],…

Optimization and Control · Mathematics 2021-04-26 Jérôme Darbon , Gabriel P. Langlois , Tingwei Meng

Training neural networks via backpropagation is often hindered by vanishing or exploding gradients. In this work, we design architectures that mitigate these issues by analyzing and controlling the network Jacobian. We first provide a…

Machine Learning · Computer Science 2026-02-12 Alex Massucco , Davide Murari , Carola-Bibiane Schönlieb

Binary Neural Networks are a promising technique for implementing efficient deep models with reduced storage and computational requirements. The training of these is however, still a compute-intensive problem that grows drastically with the…

Neural network training is typically viewed as gradient descent on a loss surface. We propose a fundamentally different perspective: learning is a structure-preserving transformation (a functor L) between the space of network parameters…

Machine Learning · Computer Science 2025-10-07 Abdulrahman Tamim

The aim of this paper is twofold. - In the setting of RCD(K,$\infty$) metric measure spaces, we derive uniform gradient and Laplacian contraction estimates along solutions of the viscous approximation of the Hamilton--Jacobi equation. We…

Probability · Mathematics 2024-09-16 Nicola Gigli , Luca Tamanini , Dario Trevisan

There has been a wave of interest in applying machine learning to study dynamical systems. We present a Hamiltonian neural network that solves the differential equations that govern dynamical systems. This is an equation-driven machine…

Computational Physics · Physics 2022-07-01 Marios Mattheakis , David Sondak , Akshunna S. Dogra , Pavlos Protopapas

We examine gradient descent on unregularized logistic regression problems, with homogeneous linear predictors on linearly separable datasets. We show the predictor converges to the direction of the max-margin (hard margin SVM) solution. The…

Machine Learning · Statistics 2024-10-29 Daniel Soudry , Elad Hoffer , Mor Shpigel Nacson , Suriya Gunasekar , Nathan Srebro

The fragility of deep neural networks to adversarially-chosen inputs has motivated the need to revisit deep learning algorithms. Including adversarial examples during training is a popular defense mechanism against adversarial attacks. This…

Optimization and Control · Mathematics 2020-05-05 Jacob H. Seidman , Mahyar Fazlyab , Victor M. Preciado , George J. Pappas

This paper analyzes the convergence and generalization of training a one-hidden-layer neural network when the input features follow the Gaussian mixture model consisting of a finite number of Gaussian distributions. Assuming the labels are…

Machine Learning · Computer Science 2023-01-30 Hongkang Li , Shuai Zhang , Meng Wang

Computing tasks may often be posed as optimization problems. The objective functions for real-world scenarios are often nonconvex and/or nondifferentiable. State-of-the-art methods for solving these problems typically only guarantee…

Optimization and Control · Mathematics 2022-10-11 Howard Heaton , Samy Wu Fung , Stanley Osher

Although deep learning has shown its powerful performance in many applications, the mathematical principles behind neural networks are still mysterious. In this paper, we consider the problem of learning a one-hidden-layer neural network…

Machine Learning · Computer Science 2019-07-17 Shuhao Xia , Yuanming Shi

First-order optimization algorithms are widely used today. Two standard building blocks in these algorithms are proximal operators (proximals) and gradients. Although gradients can be computed for a wide array of functions, explicit…

Optimization and Control · Mathematics 2023-05-30 Stanley Osher , Howard Heaton , Samy Wu Fung
‹ Prev 1 3 4 5 6 7 10 Next ›