中文
相关论文

相关论文: Globally injective and bijective neural operators

200 篇论文

Neural networks are a powerful class of functions that can be trained with simple gradient descent to achieve state-of-the-art performance on a variety of applications. Despite their practical success, there is a paucity of results that…

机器学习 · 计算机科学 2017-03-06 Bo Xie , Yingyu Liang , Le Song

While deep learning has outperformed other methods for various tasks, theoretical frameworks that explain its reason have not been fully established. To address this issue, we investigate the excess risk of two-layer ReLU neural networks in…

机器学习 · 统计学 2022-06-07 Shunta Akiyama , Taiji Suzuki

Any Lipschitz map $f\colon M \to N$ between metric spaces can be "linearised" in such a way that it becomes a bounded linear operator $\widehat{f}\colon \mathcal F(M) \to \mathcal F(N)$ between the Lipschitz-free spaces over $M$ and $N$.…

泛函分析 · 数学 2022-12-15 Luis García-Lirola , Colin Petitjean , Antonin Prochazka

The study of the expressive power of neural networks has investigated the fundamental limits of neural networks. Most existing results assume real-valued inputs and parameters as well as exact operations during the evaluation of neural…

机器学习 · 计算机科学 2024-07-17 Yeachan Park , Geonho Hwang , Wonyeol Lee , Sejun Park

Linear algebra's main concerns are sets of vectors, linear functions, subspaces, linear systems, matrices and concepts about those, such as whether the solution of linear system exists or is unique; a set of vectors is linearly independent…

符号计算 · 计算机科学 2025-04-15 Iago Leal de Freitas , Júlia Mota , João Paixão , Lucas Rufino

In the first half of this text we explore the interrelationships between the abstract theory of limit operators (see e.g. the recent monographs of Rabinovich, Roch & Silbermann and Lindner) and the concepts and results of the generalised…

谱理论 · 数学 2010-11-05 Simon N. Chandler-Wilde , Marko Lindner

Several non-linear operators in stochastic analysis, such as solution maps to stochastic differential equations, depend on a temporal structure which is not leveraged by contemporary neural operators designed to approximate general maps…

动力系统 · 数学 2025-04-11 Luca Galimberti , Anastasis Kratsios , Giulia Livieri

Equivariant neural networks have shown improved performance, expressiveness and sample complexity on symmetrical domains. But for some specific symmetries, representations, and choice of coordinates, the most common point-wise activations,…

机器学习 · 计算机科学 2024-01-18 Marco Pacini , Xiaowen Dong , Bruno Lepri , Gabriele Santin

Deep learning models are often successfully trained using gradient descent, despite the worst case hardness of the underlying non-convex optimization problem. The key question is then under what conditions can one prove that optimization…

机器学习 · 计算机科学 2017-02-28 Alon Brutzkus , Amir Globerson

Computationally efficient surrogates for parametrized physical models play a crucial role in science and engineering. Operator learning provides data-driven surrogates that map between function spaces. However, instead of full-field…

机器学习 · 计算机科学 2024-12-31 Daniel Zhengyu Huang , Nicholas H. Nelsen , Margaret Trautner

In this paper, we have extended the well-established universal approximator theory to neural networks that use the unbounded ReLU activation function and a nonlinear softmax output layer. We have proved that a sufficiently large neural…

机器学习 · 计算机科学 2020-02-12 Behnam Asadi , Hui Jiang

Navigation is crucial for animal behavior and is assumed to require an internal representation of the external environment, termed a cognitive map. The precise form of this representation is often considered to be a metric representation of…

神经元与认知 · 定量生物学 2020-02-10 Tie Xu , Omri Barak

Deep neural networks, particularly those employing Rectified Linear Units (ReLU), are often perceived as complex, high-dimensional, non-linear systems. This complexity poses a significant challenge to understanding their internal learning…

机器学习 · 计算机科学 2025-11-11 Longqing Ye

Amongst others, the adoption of Rectified Linear Units (ReLUs) is regarded as one of the ingredients of the success of deep learning. ReLU activation has been shown to mitigate the vanishing gradient issue, to encourage sparsity in the…

机器学习 · 统计学 2021-10-14 Nicola Picchiotti , Marco Gori

Convex functions and their gradients play a critical role in mathematical imaging, from proximal optimization to Optimal Transport. The successes of deep learning has led many to use learning-based methods, where fixed functions or…

机器学习 · 计算机科学 2025-04-09 Anne Gagneux , Mathurin Massias , Emmanuel Soubies , Rémi Gribonval

We prove that a $C^{\infty}$ semialgebraic local diffeomorphism of $\mathbb{R}^n$ with non-properness set having codimension greater than or equal to $2$ is a global diffeomorphism if $n-1$ suitable linear partial differential operators are…

几何拓扑 · 数学 2024-04-30 Francisco Braun , Luis Renato Gonçalves Dias , Jean Venato Santos

Understanding and mapping a new environment are core abilities of any autonomously navigating agent. While classical robotics usually estimates maps in a stand-alone manner with SLAM variants, which maintain a topological or metric…

计算机视觉与模式识别 · 计算机科学 2023-09-28 Pierre Marza , Laetitia Matignon , Olivier Simonin , Christian Wolf

This paper ascertains the global behavior of the forward and backward branches of solutions provided by the Leray-Schauder continuation theorem for orientable $\mathcal{C}^1$ Fredholm maps, as developed by the authors in [54]. Under…

偏微分方程分析 · 数学 2025-12-10 Julián López-Gómez , Juan Carlos Sampedro

Neural networks are universal function approximators which are known to generalize well despite being dramatically overparameterized. We study this phenomenon from the point of view of the spectral bias of neural networks. Our contributions…

机器学习 · 计算机科学 2022-09-07 Qingguo Hong , Jonathan W. Siegel , Qinyang Tan , Jinchao Xu

The inductive biases of trained neural networks are difficult to understand and, consequently, to adapt to new settings. We study the inductive biases of linearizations of neural networks, which we show to be surprisingly good summaries of…