English
Related papers

Related papers: Tensor Programs III: Neural Matrix Laws

200 papers

We consider operator-valued polynomials in Gaussian Unitary Ensemble random matrices and we show that its $L^p$-norm can be upper bounded, up to an asymptotically small error, by the operator norm of the same polynomial evaluated in free…

Probability · Mathematics 2024-10-31 Félix Parraud

The paper discusses the capabilities of multilayer perceptron neural networks implementing metric recognition methods, for which the values of the weights are calculated analytically by formulas. Comparative experiments in training a neural…

Machine Learning · Computer Science 2025-06-02 Polad Geidarov

This work presents an approach to automatically induction for non-greedy decision trees constructed from neural network architecture. This construction can be used to transfer weights when growing or pruning a decision tree, allowing…

Machine Learning · Statistics 2019-12-10 Chapman Siu

Many complex systems--from social and communication networks to biological networks and the Internet--are thought to exhibit scale-free structure. However, prevailing explanations rely on the constant addition of new nodes, an assumption…

Adaptation and Self-Organizing Systems · Physics 2022-11-10 Christopher W. Lynn , Caroline M. Holmes , Stephanie E. Palmer

Backward propagation (BP) is widely used to compute the gradients in neural network training. However, it is hard to implement BP on edge devices due to the lack of hardware and software resources to support automatic differentiation. This…

Machine Learning · Computer Science 2023-10-11 Yequan Zhao , Xinling Yu , Zhixiong Chen , Ziyue Liu , Sijia Liu , Zheng Zhang

In this paper, we investigate a two-layer fully connected neural network of the form $f(X)=\frac{1}{\sqrt{d_1}}\boldsymbol{a}^\top \sigma\left(WX\right)$, where $X\in\mathbb{R}^{d_0\times n}$ is a deterministic data matrix,…

Statistics Theory · Mathematics 2023-04-17 Zhichao Wang , Yizhe Zhu

We study random matrices acting on tensor product spaces which have been transformed by a linear block operation. Using operator-valued free probability theory, under some mild assumptions on the linear map acting on the blocks, we compute…

Probability · Mathematics 2016-01-26 Octavio Arizmendi , Ion Nechita , Carlos Vargas

We investigate quantum algorithms derived from tensor networks to simulate the static and dynamic properties of quantum many-body systems. Using a sequentially prepared quantum circuit representation of a matrix product state (MPS) that we…

Quantum Physics · Physics 2024-12-04 Michael L. Wall , Aidan Reilly , John S. Van Dyke , Collin Broholm , Paraj Titum

Physics-Informed Neural Networks (PINNs) have emerged recently as a promising application of deep neural networks to the numerical solution of nonlinear partial differential equations (PDEs). However, it has been recognized that adaptive…

Machine Learning · Computer Science 2024-06-21 Levi McClenny , Ulisses Braga-Neto

The paper deals with distribution of singular values of product of random matrices arising in the analysis of deep neural networks. The matrices resemble the product analogs of the sample covariance matrices, however, an important…

Mathematical Physics · Physics 2020-11-23 Leonid Pastur

The deep learning literature is continuously updated with new architectures and training techniques. However, weight initialization is overlooked by most recent research, despite some intriguing findings regarding random weights. On the…

Neural and Evolutionary Computing · Computer Science 2022-07-19 Leonardo Scabini , Bernard De Baets , Odemir M. Bruno

Biological neurons exhibit remarkable intelligence: they maintain internal states, communicate selectively with other neurons, and self-organize into complex graphs rather than rigid hierarchical layers. What if artificial intelligence…

Machine Learning · Computer Science 2025-12-01 Antoine Salomon

Network science can offer fundamental insights into the structural and functional properties of complex systems. For example, it is widely known that neuronal circuits tend to organize into basic functional topological modules, called…

Adaptation and Self-Organizing Systems · Physics 2022-08-03 Matteo Zambra , Alberto Testolin , Amos Maritan

We provide several new results on the sample complexity of vector-valued linear predictors (parameterized by a matrix), and more generally neural networks. Focusing on size-independent bounds, where only the Frobenius norm distance of the…

Machine Learning · Computer Science 2023-10-26 Roey Magen , Ohad Shamir

This work proposes a Neural Network model that can control its depth using an iterate-to-fixed-point operator. The architecture starts with a standard layered Network but with added connections from current later to earlier layers, along…

Machine Learning · Computer Science 2021-11-02 Mansura Habiba , Barak A. Pearlmutter

We present the surprising result that randomly initialized neural networks are good feature extractors in expectation. These random features correspond to finite-sample realizations of what we call Neural Network Prior Kernel (NNPK), which…

Machine Learning · Computer Science 2022-02-15 Ehsan Amid , Rohan Anil , Wojciech Kotłowski , Manfred K. Warmuth

We introduce a mixed integer program (MIP) for assigning importance scores to each neuron in deep neural network architectures which is guided by the impact of their simultaneous pruning on the main learning task of the network. By…

Machine Learning · Computer Science 2023-07-26 Mostafa ElAraby , Guy Wolf , Margarida Carvalho

This paper is on improving the training of binary neural networks in which both activations and weights are binary. While prior methods for neural network binarization binarize each filter independently, we propose to instead parametrize…

Computer Vision and Pattern Recognition · Computer Science 2019-04-17 Adrian Bulat , Jean Kossaifi , Georgios Tzimiropoulos , Maja Pantic

Many state-of-the-art results obtained with deep networks are achieved with the largest models that could be trained, and if more computation power was available, we might be able to exploit much larger datasets in order to improve…

Machine Learning · Statistics 2014-07-01 Kyunghyun Cho , Yoshua Bengio

Leaky integrate-and-fire (LIF) models are mean-field limits, with a large number of neurons, used to describe neural networks. We consider inhomogeneous networks structured by a connec-tivity parameter (strengths of the synaptic weights)…

Neurons and Cognition · Quantitative Biology 2017-06-20 Benoît Perthame , Delphine Salort , Gilles Wainrib