中文
相关论文

相关论文: A Capacity Scaling Law for Artificial Neural Netwo…

200 篇论文

The aim of this thesis is to compare the capacity of different models of neural networks. We start by analysing the problem solving capacity of a single perceptron using a simple combinatorial argument. After some observations on the…

无序系统与神经网络 · 物理学 2022-11-15 Leonardo Cruciani

Memorization is worst-case generalization. Based on MacKay's information theoretic model of supervised machine learning, this article discusses how to practically estimate the maximum size of a neural network given a training data set.…

神经与进化计算 · 计算机科学 2018-10-05 Gerald Friedland , Alfredo Metere , Mario Krell

We use a formal correspondence between thermodynamics and inference, where the number of samples can be thought of as the inverse temperature, to study a quantity called ``learning capacity'' which is a measure of the effective…

机器学习 · 计算机科学 2024-10-22 Daiwei Chen , Wei-Kai Chang , Pratik Chaudhari

We define a neural network as a septuple consisting of (1) a state vector, (2) an input projection, (3) an output projection, (4) a weight matrix, (5) a bias vector, (6) an activation map and (7) a loss function. We argue that the loss…

机器学习 · 计算机科学 2021-02-15 Vitaly Vanchurin

We study the capacity of \emph{sign} perceptrons neural networks (SPNN) and particularly focus on 1-hidden layer \emph{treelike committee machine} (TCM) architectures. Similarly to what happens in the case of a single perceptron neuron, it…

无序系统与神经网络 · 物理学 2023-12-14 Mihailo Stojnic

Capacity analysis has been recently introduced as a way to analyze how linear models distribute their modelling capacity across the input space. In this paper, we extend the notion of capacity allocation to the case of neural networks with…

机器学习 · 计算机科学 2019-02-28 Jonathan Donier

Existing scaling laws for Large Language Models (LLMs), predominantly monotonic power laws, fail to explain emerging non-monotonic phenomena such as catastrophic overtraining and quantization-induced degradation, where performance…

机器学习 · 计算机科学 2026-05-25 Xu Ouyang , Deyi Liu , Yuhang Cai , Jing Liu , Yuan Yang , Chen Zheng , Thomas Hartvigsen , Yiyuan Ma

We investigate the VC-dimension of the perceptron and simple two-layer networks like the committee- and the parity-machine with weights restricted to values $\pm1$. For binary inputs, the VC-dimension is determined by atypical pattern sets,…

凝聚态物理 · 物理学 2009-10-28 S. Mertens , A. Engel

Understanding the memory capacity of neural networks remains a challenging problem in implementing artificial intelligence systems. In this paper, we address the notion of capacity with respect to Hopfield networks and propose a dynamic…

神经与进化计算 · 计算机科学 2017-09-19 Saarthak Sarup , Mingoo Seok

We consider the problem of estimating an upper bound on the capacity of a memoryless channel with unknown channel law and continuous output alphabet. A novel data-driven algorithm is proposed that exploits the dual representation of…

信息论 · 计算机科学 2024-01-25 Christian Häger , Erik Agrell

In this paper, we characterize the information-theoretic capacity scaling of wireless ad hoc networks with $n$ randomly distributed nodes. By using an exact channel model from Maxwell's equations, we successfully resolve the conflict in the…

信息论 · 计算机科学 2011-10-19 Si-Hyeon Lee , Sae-Young Chung

We consider the memorization capabilities of multilayered \emph{sign} perceptrons neural networks (SPNNs). A recent rigorous upper-bounding capacity characterization, obtained in \cite{Stojnictcmspnncaprdt23} utilizing the Random Duality…

机器学习 · 统计学 2023-12-14 Mihailo Stojnic

A long standing open problem in the theory of neural networks is the development of quantitative methods to estimate and compare the capabilities of different architectures. Here we define the capacity of an architecture by the binary…

机器学习 · 计算机科学 2019-03-29 Pierre Baldi , Roman Vershynin

Graph representation learning has become a standard approach for analyzing networked data, with latent embeddings widely used for link prediction, community detection, and related tasks. Yet a basic design choice, the latent dimension, is…

Determining the memory capacity of two layer neural networks with $m$ hidden neurons and input dimension $d$ (i.e., $md+2m$ total trainable parameters), which refers to the largest size of general data the network can memorize, is a…

机器学习 · 计算机科学 2024-07-25 Liam Madden , Christos Thrampoulidis

We study the fundamental network capacity of a multi-user wireless downlink under two assumptions: (1) Channels are not explicitly measured and thus instantaneous states are unknown, (2) Channels are modeled as ON/OFF Markov chains. This is…

信息论 · 计算机科学 2010-04-23 Chih-ping Li , Michael J. Neely

Understanding the theoretical foundations of how memories are encoded and retrieved in neural populations is a central challenge in neuroscience. A popular theoretical scenario for modeling memory function is the attractor neural network…

神经元与认知 · 定量生物学 2016-02-17 Alireza Alemi , Carlo Baldassi , Nicolas Brunel , Riccardo Zecchina

The notion of memory capacity, originally introduced for echo state and linear networks with independent inputs, is generalized to nonlinear recurrent networks with stationary but dependent inputs. The presence of dependence in the inputs…

最优化与控制 · 数学 2020-10-28 Lukas Gonon , Lyudmila Grigoryeva , Juan-Pablo Ortega

On a variety of tasks, the performance of neural networks predictably improves with training time, dataset size and model size across many orders of magnitude. This phenomenon is known as a neural scaling law. Of fundamental importance is…

机器学习 · 统计学 2024-06-25 Blake Bordelon , Alexander Atanasov , Cengiz Pehlevan

Deep neural networks are highly expressive machine learning models with the ability to interpolate arbitrary datasets. Deep nets are typically optimized via first-order methods and the optimization process crucially depends on the…

机器学习 · 统计学 2019-11-12 Talha Cihad Gulcu
‹ 上一页 1 2 3 10 下一页 ›