中文
相关论文

相关论文: Numerical Analysis on Neural Network Projected Sch…

200 篇论文

Gaussian mixture models form a flexible and expressive parametric family of distributions that has found applications in a wide variety of applications. Unfortunately, fitting these models to data is a notoriously hard problem from a…

统计理论 · 数学 2023-01-05 Yuling Yan , Kaizheng Wang , Philippe Rigollet

We describe tests validating progress made toward acceleration and automation of hydrodynamic codes in the regime of developed turbulence by three Deep Learning (DL) Neural Network (NN) schemes trained on Direct Numerical Simulations of…

流体动力学 · 物理学 2018-12-06 Ryan King , Oliver Hennigh , Arvind Mohan , Michael Chertkov

We present a novel approach to approximate Gaussian and mixture-of-Gaussians filtering. Our method relies on a variational approximation via a gradient-flow representation. The gradient flow is derived from a Kullback--Leibler discrepancy…

统计计算 · 统计学 2023-06-21 Adrien Corenflos , Hany Abdulsamad

A direct numerical simulation (DNS) of a channel flow with one curved surface was performed at moderate Reynolds number (Re_tau = 395 at the inlet). The adverse pressure gradient was obtained by a wall curvature through a mathematical…

流体动力学 · 物理学 2017-11-22 Matthieu Marquillie , Jean-Philippe Laval , Rostislav Dolganov

Implicit deep learning has received increasing attention recently due to the fact that it generalizes the recursive prediction rules of many commonly used neural network architectures. Its prediction rule is provided implicitly based on the…

机器学习 · 计算机科学 2022-02-21 Tianxiang Gao , Hailiang Liu , Jia Liu , Hridesh Rajan , Hongyang Gao

We introduce a (de)-regularization of the Maximum Mean Discrepancy (DrMMD) and its Wasserstein gradient flow. Existing gradient flows that transport samples from source distribution to target distribution with only target samples, either…

A popular method to perform adversarial attacks on neuronal networks is the so-called fast gradient sign method and its iterative variant. In this paper, we interpret this method as an explicit Euler discretization of a differential…

机器学习 · 计算机科学 2025-09-17 Lukas Weigand , Tim Roith , Martin Burger

We propose a probabilistic way for reducing the cost of classical projection-based model order reduction methods for parameter-dependent linear equations. A reduced order model is here approximated from its random sketch, which is a set of…

数值分析 · 数学 2020-05-19 Oleg Balabanov , Anthony Nouy

The mean-field theory for two-layer neural networks considers infinitely wide networks that are linearly parameterized by a probability measure over the parameter space. This nonparametric perspective has significantly advanced both the…

机器学习 · 计算机科学 2025-08-08 Sinho Chewi , Philippe Rigollet , Yuling Yan

The Wasserstein space of probability measures is known for its intricate Riemannian structure, which underpins the Wasserstein geometry and enables gradient flow algorithms. However, the Wasserstein geometry may not be suitable for certain…

偏微分方程分析 · 数学 2025-05-23 Zhengxin Zhang , Ziv Goldfeld , Kristjan Greenewald , Youssef Mroueh , Bharath K. Sriperumbudur

Recent results have shown that for two-layer fully connected neural networks, gradient flow converges to a global optimum in the infinite width limit, by making a connection between the mean field dynamics and the Wasserstein gradient flow.…

最优化与控制 · 数学 2020-07-16 Walid Krichene , Kenneth F. Caluya , Abhishek Halder

Many score-based active learning methods have been successfully applied to graph-structured data, aiming to reduce the number of labels and achieve better performance of graph neural networks based on predefined score functions. However,…

机器学习 · 计算机科学 2023-04-25 Yinchuan Li , Zhigang Li , Wenqian Li , Yunfeng Shao , Yan Zheng , Jianye Hao

We give a comprehensive description of Wasserstein gradient flows of maximum mean discrepancy (MMD) functionals $\mathcal F_\nu := \text{MMD}_K^2(\cdot, \nu)$ towards given target measures $\nu$ on the real line, where we focus on the…

偏微分方程分析 · 数学 2025-12-05 Richard Duong , Viktor Stein , Robert Beinert , Johannes Hertrich , Gabriele Steidl

We reveal a precise mathematical framework about a new family of generative models which we call Gradient Flow Drifting. With this framework, we prove an equivalence between the recently proposed Drifting Model and the Wasserstein gradient…

机器学习 · 计算机科学 2026-03-12 Jiarui Cao , Zixuan Wei , Yuxin Liu

We show that the continuous-time gradient descent in Rn can be viewed as an optimal controlled evolution for a suitable action functional; a similar result holds for stochastic gradient descent. We then provide an analogous characterization…

最优化与控制 · 数学 2025-11-03 Yongxin Chen , Tryphon Georgiou , Michele Pavon

Natural gradient descent is a principled method for adapting the parameters of a statistical model on-line using an underlying Riemannian parameter space to redefine the direction of steepest descent. The algorithm is examined via methods…

无序系统与神经网络 · 物理学 2009-10-31 Magnus Rattray , David Saad

In the field of fluid numerical analysis, there has been a long-standing problem: lacking of a rigorous mathematical tool to map from a continuous flow field to discrete vortex particles, hurdling the Lagrangian particles from inheriting…

计算物理 · 物理学 2023-09-14 Shiying Xiong , Xingzhe He , Yunjin Tong , Yitong Deng , Bo Zhu

Neural networks trained with standard objectives exhibit behaviors characteristic of probabilistic inference: soft clustering, prototype specialization, and Bayesian uncertainty tracking. These phenomena appear across architectures -- in…

机器学习 · 计算机科学 2026-01-01 Alan Oursland

We derive explicit equations governing the cumulative biases and weights in Deep Learning with ReLU activation function, based on gradient descent for the Euclidean cost in the input layer, and under the assumption that the weights are, in…

机器学习 · 计算机科学 2025-01-15 Thomas Chen

We investigate gradient descent training of wide neural networks and the corresponding implicit bias in function space. For univariate regression, we show that the solution of training a width-$n$ shallow ReLU network is within $n^{- 1/2}$…

机器学习 · 统计学 2023-05-30 Hui Jin , Guido Montúfar