English
Related papers

Related papers: Deep BSDE Solver on Bounded Domains Part I: Genera…

200 papers

We propose new machine learning schemes for solving high dimensional nonlinear partial differential equations (PDEs). Relying on the classical backward stochastic differential equation (BSDE) representation of PDEs, our algorithms estimate…

Probability · Mathematics 2020-06-08 Côme Huré , Huyên Pham , Xavier Warin

We propose a new algorithm for solving parabolic partial differential equations (PDEs) and backward stochastic differential equations (BSDEs) in high dimension, by making an analogy between the BSDE and reinforcement learning with the…

Numerical Analysis · Mathematics 2020-07-14 Weinan E , Jiequn Han , Arnulf Jentzen

In this paper, we study a multidimensional backward stochastic differential equation (BSDE) with an additional rough drift (rough BSDE), and give the existence and uniqueness of the adapted solution, either when the terminal value and the…

Probability · Mathematics 2024-01-12 Jiahao Liang , Shanjian Tang

This paper addresses the general problem of domain adaptation which arises in a variety of applications where the distribution of the labeled sample available somewhat differs from that of the test data. Building on previous work by…

Machine Learning · Computer Science 2023-12-04 Yishay Mansour , Mehryar Mohri , Afshin Rostamizadeh

This paper establishes risk convergence and asymptotic weight matrix alignment --- a form of implicit regularization --- of gradient flow and gradient descent when applied to deep linear networks on linearly separable data. In more detail,…

Machine Learning · Computer Science 2019-02-26 Ziwei Ji , Matus Telgarsky

Deep learning has been shown to achieve impressive results in several domains like computer vision and natural language processing. A key element of this success has been the development of new loss functions, like the popular cross-entropy…

Machine Learning · Computer Science 2019-07-19 Francesco Giannini , Giuseppe Marra , Michelangelo Diligenti , Marco Maggini , Marco Gori

We present a convergence rate analysis for biased stochastic gradient descent (SGD), where individual gradient updates are corrupted by computation errors. We develop stochastic quadratic constraints to formulate a small linear matrix…

Optimization and Control · Mathematics 2020-03-31 Bin Hu , Peter Seiler , Laurent Lessard

We establish the rectifiability of measures satisfying a linear PDE constraint. The obtained rectifiability dimensions are optimal for many usual PDE operators, including all first-order systems and all second-order scalar operators. In…

Analysis of PDEs · Mathematics 2019-02-01 Adolfo Arroyo-Rabasa , Guido De Philippis , Jonas Hirsch , Filip Rindler

Diffusion approximation provides weak approximation for stochastic gradient descent algorithms in a finite time horizon. In this paper, we introduce new tools motivated by the backward error analysis of numerical stochastic differential…

Machine Learning · Computer Science 2019-09-05 Yuanyuan Feng , Tingran Gao , Lei Li , Jian-Guo Liu , Yulong Lu

Machine learning for partial differential equations (PDEs) is a hot topic. In this paper we introduce and analyse a Deep BSDE scheme for nonlinear integro-PDEs with unbounded nonlocal operators -problems arising in e.g. stochastic control…

Analysis of PDEs · Mathematics 2024-07-15 Espen Robstad Jakobsen , Sehail Mazid

Generalization error bounds for deep neural networks trained by stochastic gradient descent (SGD) are derived by combining a dynamical control of an appropriate parameter norm and the Rademacher complexity estimate based on parameter norms.…

Machine Learning · Computer Science 2023-05-30 Mingze Wang , Chao Ma

In this work we propose a new algorithm for solving high-dimensional backward stochastic differential equations (BSDEs). Based on the general theta-discretization for the time-integrands, we show how to efficiently use eXtreme Gradient…

Numerical Analysis · Mathematics 2021-07-15 Long Teng

In this paper, we propose a new covering technique localized for the trajectories of SGD. This localization provides an algorithm-specific complexity measured by the covering number, which can have dimension-independent cardinality in…

Machine Learning · Statistics 2022-09-20 Sejun Park , Umut Şimşekli , Murat A. Erdogdu

In this work, we describe a generic approach to show convergence with high probability for stochastic convex optimization. In previous works, either the convergence is only in expectation or the bound depends on the diameter of the domain.…

Optimization and Control · Mathematics 2022-10-04 Alina Ene , Huy L. Nguyen

By establishing a local version of Bismut formula for Dirichlet semigroups on a regular domain, gradient estimates are derived for killed SDEs with singular drifts. As an application, the total variation distance between two solutions of…

Probability · Mathematics 2026-03-30 Feng-Yu Wang , Xiao-Yu Zhao

The loss function is crucial to machine learning, especially in supervised learning frameworks. It is a fundamental component that controls the behavior and general efficacy of learning algorithms. However, despite their widespread use,…

Machine Learning · Computer Science 2026-02-09 Soumi Mahato , Lineesh M. C

We provide a unified framework that applies to a general family of convex losses across binary and multiclass settings in the overparameterized regime to approximately characterize the implicit bias of gradient descent in closed form.…

Machine Learning · Statistics 2025-06-11 Kuo-Wei Lai , Vidya Muthukumar

Stochastic Gradient Descent (SGD) based methods have been widely used for training large-scale machine learning models that also generalize well in practice. Several explanations have been offered for this generalization performance, a…

Machine Learning · Computer Science 2021-02-11 Yikai Zhang , Wenjia Zhang , Sammy Bald , Vamsi Pingali , Chao Chen , Mayank Goswami

Uniform deviation bounds limit the difference between a model's expected loss and its loss on an empirical sample uniformly for all models in a learning problem. As such, they are a critical component to empirical risk minimization. In this…

Machine Learning · Statistics 2017-02-28 Olivier Bachem , Mario Lucic , S. Hamed Hassani , Andreas Krause

In many Bayesian inverse problems the goal is to recover a spatially varying random field. Such problems are often computationally challenging especially when the forward model is governed by complex partial differential equations (PDEs).…

Numerical Analysis · Mathematics 2022-11-09 Zhihang Xu , Qifeng Liao , Jinglai Li