中文
相关论文

相关论文: Function-Space Optimality of Neural Architectures …

200 篇论文

Humans and animals can recognize latent structures in their environment and apply this information to efficiently navigate the world. However, it remains unclear what aspects of neural activity contribute to these computational…

神经元与认知 · 定量生物学 2024-04-12 Albert J. Wakhloo , Will Slatton , SueYeon Chung

The aim of this paper is to present a survey of some recent results obtained in the study of spaces with asymmetric norm. The presentation follows the ideas from the theory of normed spaces (topology, continuous linear operators, continuous…

泛函分析 · 数学 2016-08-14 S. Cobzaş

In this note, we consider the smallest submaximal space structure {\mu}(X) on a Banach space X. We derive a characterization of {\mu}(X) up to complete isometric isomorphism in terms of a universal property. Also, we show that an injective…

算子代数 · 数学 2012-12-12 Vinod Kumar P. , M. S. Balasubramani

Activation functions are non-linearities in neural networks that allow them to learn complex mapping between inputs and outputs. Typical choices for activation functions are ReLU, Tanh, Sigmoid etc., where the choice generally depends on…

We study neural field equations, which are prototypical models of large-scale cortical activity, subject to random data. We view this spatially-extended, nonlocal evolution equation as a Cauchy problem on abstract Banach spaces, with…

Many modern neural network architectures are trained in an overparameterized regime where the parameters of the model exceed the size of the training dataset. Sufficiently overparameterized neural network architectures in principle have the…

机器学习 · 计算机科学 2019-02-14 Samet Oymak , Mahdi Soltanolkotabi

A long standing open problem in the theory of neural networks is the development of quantitative methods to estimate and compare the capabilities of different architectures. Here we define the capacity of an architecture by the binary…

机器学习 · 计算机科学 2019-03-29 Pierre Baldi , Roman Vershynin

Function-space priors in Bayesian Neural Networks (BNNs) provide a more intuitive approach to embedding beliefs directly into the model's output, thereby enhancing regularization, uncertainty quantification, and risk-aware decision-making.…

机器学习 · 计算机科学 2025-08-13 Marcin Sendera , Amin Sorkhei , Tomasz Kuśmierczyk

The idea of best approximation in linear n-normed space is presented and some examples showing various possibilities of best approximations in linear n-normed space is given. Also, we study strictly convex n-norm and enquire about the…

泛函分析 · 数学 2023-09-27 Prasenjit Ghosh , T. K. Samanta

Designing neural networks typically relies on manual trial and error or a neural architecture search (NAS) followed by weight training. The former is time-consuming and labor-intensive, while the latter often discretizes architecture search…

机器学习 · 计算机科学 2025-11-19 Zitong Huang , Mansooreh Montazerin , Ajitesh Srivastava

The training process of neural networks usually optimize weights and bias parameters of linear transformations, while nonlinear activation functions are pre-specified and fixed. This work develops a systematic approach to constructing…

机器学习 · 计算机科学 2024-10-29 Zhengqi Liu , Shuhao Cao , Yuwen Li , Ludmil Zikatanov

We study the expressive power of deep ReLU neural networks for approximating functions in dilated shift-invariant spaces, which are widely used in signal processing, image processing, communications and so on. Approximation error bounds are…

机器学习 · 计算机科学 2023-12-05 Yunfei Yang , Zhen Li , Yang Wang

Researchers commonly believe that neural networks model a high-dimensional space but cannot give a clear definition of this space. What is this space? What is its dimension? And does it has finite dimensions? In this paper, we develop a…

机器学习 · 计算机科学 2023-05-10 John Chiang

We study regularized deep neural networks (DNNs) and introduce a convex analytic framework to characterize the structure of the hidden layers. We show that a set of optimal hidden layer weights for a norm regularized DNN training problem…

机器学习 · 计算机科学 2021-06-14 Tolga Ergen , Mert Pilanci

Graph convolutional neural network (GCNN) operates on graph domain and it has achieved a superior performance to accomplish a wide range of tasks. In this paper, we introduce a Barron space of functions on a compact domain of graph signals.…

机器学习 · 统计学 2023-11-07 Seok-Young Chung , Qiyu Sun

Neural networks have to capture mathematical relationships in order to learn various tasks. They approximate these relations implicitly and therefore often do not generalize well. The recently proposed Neural Arithmetic Logic Unit (NALU) is…

神经与进化计算 · 计算机科学 2020-03-18 Daniel Schlör , Markus Ring , Andreas Hotho

This paper studies Tikhonov regularization for finitely smoothing operators in Banach spaces when the penalization enforces too much smoothness in the sense that the penalty term is not finite at the true solution. In a Hilbert space…

数值分析 · 数学 2022-03-07 Philip Miller , Thorsten Hohage

Foundation models have revolutionized various fields such as natural language processing (NLP) and computer vision (CV). While efforts have been made to transfer the success of the foundation models in general AI domains to biology,…

机器学习 · 计算机科学 2026-05-29 Yi Fang , Haoran Xu , Jiaxin Han , Sirui Ding , Yizhi Wang , Yue Wang , Xuan Wang

Deep neural networks owe their expressive power to nonlinear activation functions. The effective field theory of signal propagation at initialization reveals a few distinct universality classes of activations that exhibit different depth…

无序系统与神经网络 · 物理学 2026-05-08 Omri Lesser , Debanjan Chowdhury

Deep learning, in the form of artificial neural networks, has achieved remarkable practical success in recent years, for a variety of difficult machine learning applications. However, a theoretical explanation for this remains a major open…

机器学习 · 计算机科学 2016-06-15 Itay Safran , Ohad Shamir