English
Related papers

Related papers: Asymptotic properties of one-layer artificial neur…

200 papers

In this work we determine a process-level Large Deviation Principle (LDP) for a model of interacting particles indexed by a lattice $\mathbb{Z}^d$. The connections are random, sparse and unscaled, so that the system converges in the large…

Probability · Mathematics 2024-10-01 James MacLaurin

Recent work on deep neural network pruning has shown there exist sparse subnetworks that achieve equal or improved accuracy, training time, and loss using fewer network parameters when compared to their dense counterparts. Orthogonal to…

Machine Learning · Computer Science 2019-12-06 Justin Cosentino , Federico Zaiter , Dan Pei , Jun Zhu

Effectively scaling up deep reinforcement learning models has proven notoriously difficult due to network pathologies during training, motivating various targeted interventions such as periodic reset and architectural advances such as layer…

Machine Learning · Computer Science 2025-06-23 Guozheng Ma , Lu Li , Zilin Wang , Li Shen , Pierre-Luc Bacon , Dacheng Tao

An artificial neural network is presented based on the idea of connections between units that are only active for a specific range of input values and zero outside that range (and so are not evaluated outside the active range). The…

Neural and Evolutionary Computing · Computer Science 2016-06-15 John Loverich

An artificial neural network architecture, parameterization networks, is proposed for simulating extrapolated dynamics beyond observed data in dynamical systems. Parameterization networks are used to ensure the long term integrity of…

Chaotic Dynamics · Physics 2019-03-21 James P. L. Tan

The bivariate distribution of degrees of adjacent vertices (degree-degree distribution) is an important network characteristic defining the statistical dependencies between degrees of adjacent vertices. We show the asymptotic degree-degree…

Probability · Mathematics 2017-01-05 Mindaugas Bloznelis

For a large class of feature maps we provide a tight asymptotic characterisation of the test error associated with learning the readout layer, in the high-dimensional limit where the input dimension, hidden layer widths, and number of…

Machine Learning · Statistics 2024-06-11 Dominik Schröder , Daniil Dmitriev , Hugo Cui , Bruno Loureiro

Asymptotic properties of random graph sequences, like occurrence of a giant component or full connectivity in Erd\H{o}s-R\'enyi graphs, are usually derived with very specific choices for defining parameters. The question arises to which…

Probability · Mathematics 2024-02-20 B. J. K. Kleijn , S. Rizzelli

In random graph models, the degree distribution of an individual node should be distinguished from the (empirical) degree distribution of the graph that records the fractions of nodes with given degree. We introduce a general framework to…

Social and Information Networks · Computer Science 2018-11-14 Siddharth Pal , Armand M. Makowski

When the parameters are independently and identically distributed (initialized) neural networks exhibit undesirable properties that emerge as the number of layers increases, e.g. a vanishing dependency on the input and a concentration on…

Machine Learning · Statistics 2020-03-03 Stefano Peluchetti , Stefano Favaro

In this paper we study the problem of learning a shallow artificial neural network that best fits a training data set. We study this problem in the over-parameterized regime where the number of observations are fewer than the number of…

Machine Learning · Computer Science 2022-08-25 Mahdi Soltanolkotabi , Adel Javanmard , Jason D. Lee

Neuromorphic networks can be described in terms of coarse-grained variables, where emergent sustained behaviours spontaneously arise if stochasticity is properly taken in account. For example it has been recently found that a directed…

Adaptation and Self-Organizing Systems · Physics 2020-01-23 Ilenia Apicella , Daniel Maria Busiello , Silvia Scarpetta , Samir Suweis

In this article we prove a Generalized Asypmtotic Equipartition Property for Networked Data Structures modelled as coloured random graphs. The main techniques in this article remains large deviation principles for suitably defined empirical…

Information Theory · Computer Science 2017-12-19 K. Doku-Amponsah

Synchronous stochastic gradient descent (SGD) is the most common method used for distributed training of deep learning models. In this algorithm, each worker shares its local gradients with others and updates the parameters using the…

Machine Learning · Computer Science 2020-09-22 Negar Foroutan Eghlidi , Martin Jaggi

The representation of functions by artificial neural networks depends on a large number of parameters in a non-linear fashion. Suitable parameters of these are found by minimizing a 'loss functional', typically by stochastic gradient…

Machine Learning · Computer Science 2021-09-16 Stephan Wojtowytsch

Sparse neural networks are effective approaches to reduce the resource requirements for the deployment of deep neural networks. Recently, the concept of adaptive sparse connectivity, has emerged to allow training sparse neural networks from…

We consider high-dimensional distribution estimation through autoregressive networks. By combining the concepts of sparsity, mixtures and parameter sharing we obtain a simple model which is fast to train and which achieves state-of-the-art…

Machine Learning · Statistics 2016-04-28 Marc Goessling , Yali Amit

In this paper, a sparsity-aware adaptive algorithm for distributed learning in diffusion networks is developed. The algorithm follows the set-theoretic estimation rationale. At each time instance and at each node of the network, a closed…

Information Theory · Computer Science 2015-06-03 Symeon Chouvardas , Konstantinos Slavakis , Yannis Kopsinis , Sergios Theodoridis

A probabilistic generative network model with $n$ nodes and $m$ overlapping layers is obtained as a superposition of $m$ mutually independent Bernoulli random graphs of varying size and strength. When $n$ and $m$ are large and of the same…

Probability · Mathematics 2021-06-01 Mindaugas Bloznelis , Joona Karjalainen , Lasse Leskelä

We consider sparse random intersection graphs with the property that the clustering coefficient does not vanish as the number of nodes tends to infinity. We find explicit asymptotic expressions for the correlation coefficient of degrees of…

Probability · Mathematics 2019-08-24 Mindaugas Bloznelis , Jerzy Jaworski , Valentas Kurauskas
‹ Prev 1 3 4 5 6 7 10 Next ›