English
Related papers

Related papers: Eigenvalue Distribution of Large Random Matrices A…

200 papers

A well-conditioned Jacobian spectrum has a vital role in preventing exploding or vanishing gradients and speeding up learning of deep neural networks. Free probability theory helps us to understand and handle the Jacobian spectrum. We…

Probability · Mathematics 2020-02-13 Tomohiro Hayase

The paper deals with distribution of singular values of product of random matrices arising in the analysis of deep neural networks. The matrices resemble the product analogs of the sample covariance matrices, however, an important…

Mathematical Physics · Physics 2020-11-23 Leonid Pastur

Recent work has shown that tight concentration of the entire spectrum of singular values of a deep network's input-output Jacobian around one at initialization can speed up learning by orders of magnitude. Therefore, to guide important…

Machine Learning · Statistics 2018-02-28 Jeffrey Pennington , Samuel S. Schoenholz , Surya Ganguli

We examine the geometry of neural network training using the Jacobian of trained network parameters with respect to their initial values. Our analysis reveals low-dimensional structure in the training process which is dependent on the input…

Machine Learning · Computer Science 2024-12-12 Nora Belrose , Adam Scherlis

We study the distribution of singular values of product of random matrices pertinent to the analysis of deep neural networks. The matrices resemble the product of the sample covariance matrices, however, an important difference is that the…

Mathematical Physics · Physics 2022-07-05 L. Pastur , V. Slavin

We derive exact analytic expressions for the distributions of eigenvalues and singular values for the product of an arbitrary number of independent rectangular Gaussian random matrices in the limit of large matrix dimensions. We show that…

Statistical Mechanics · Physics 2013-05-29 Z. Burda , A. Jarosz , G. Livan , M. A. Nowak , A. Swiech

We consider fully connected and feedforward deep neural networks with dependent and possibly heavy-tailed weights, as introduced in [26], to address limitations of the standard Gaussian prior. It has been proved in [26] that, as the number…

Machine Learning · Statistics 2026-05-14 Nicola Apollonio , Giovanni Franzina , Giovanni Luca Torrisi

Complex Hermitian random matrices with a unitary symmetry can be distinguished by a weight function. When this is even, it is a known result that the distribution of the singular values can be decomposed as the superposition of two…

Probability · Mathematics 2015-03-26 Folkmar Bornemann , Peter J. Forrester

Deep neural networks are known to suffer from exploding or vanishing gradients as depth increases, a phenomenon closely tied to the spectral behavior of the input-output Jacobian. Prior work has identified critical initialization schemes…

Machine Learning · Computer Science 2025-11-25 Benjamin Dadoun , Soufiane Hayou , Hanan Salam , Mohamed El Amine Seddik , Pierre Youssef

The generalization error of deep neural networks via their classification margin is studied in this work. Our approach is based on the Jacobian matrix of a deep neural network and can be applied to networks with arbitrary non-linearities…

Machine Learning · Statistics 2017-07-04 Jure Sokolic , Raja Giryes , Guillermo Sapiro , Miguel R. D. Rodrigues

The distribution of the ratios of consecutive eigenvalue spacings of random matrices has emerged as an important tool to study spectral properties of many-body systems. This article numerically investigates the eigenvalue ratios…

Disordered Systems and Neural Networks · Physics 2022-07-13 Ankit Mishra , Tanu Raghav , Sarika Jalan

We investigate the distribution of eigenvalues of weighted adjacency matrices from a specific ensemble of random graphs. We distribute $N$ vertices across a fixed number $\kappa$ of components, with asymptotically $\alpha_j \dot N$ vertices…

Mathematical Physics · Physics 2024-09-30 Valentin Vengerovsky

We analyse the limiting behavior of the eigenvalue and singular value distribution for random convolution operators on large (not necessarily Abelian) groups, extending the results by M. Meckes for the Abelian case. We show that for regular…

Probability · Mathematics 2017-12-21 Radosław Adamczak

We study the deformation of the input space by a trained autoencoder via the Jacobians of the trained weight matrices. In doing so, we prove bounds for the mean squared errors for points in the input space, under assumptions regarding the…

Machine Learning · Computer Science 2021-07-15 Susama Agarwala , Benjamin Dees , Andrew Gearhart , Corey Lowman

Given a collection $\{\lambda_1, \dots, \lambda_n\} $ of real numbers, there is a canonical probability distribution on the set of real symmetric or complex Hermitian matrices with eigenvalues $\lambda_1,\ldots,\lambda_n$. In this paper, we…

Probability · Mathematics 2023-11-30 Elizabeth S. Meckes , Mark W. Meckes

Learning expressive probabilistic models correctly describing the data is a ubiquitous problem in machine learning. A popular approach for solving it is mapping the observations into a representation space with a simple joint distribution,…

Machine Learning · Statistics 2020-10-28 Luigi Gresele , Giancarlo Fissore , Adrián Javaloy , Bernhard Schölkopf , Aapo Hyvärinen

Motivated by recent results in random matrix theory we will study the distributions arising from products of complex Gaussian random matrices and truncations of Haar distributed unitary matrices. We introduce an appropriately general class…

Classical Analysis and ODEs · Mathematics 2014-08-28 Wolfgang Gawronski , Thorsten Neuschel , Dries Stivigny

For two large matrices ${\mathbf X}$ and ${\mathbf Y}$ with Gaussian i.i.d.\ entries and dimensions $T\times N_X$ and $T\times N_Y$, respectively, we derive the probability distribution of the singular values of $\mathbf{X}^T \mathbf{Y}$ in…

Statistics Theory · Mathematics 2025-08-29 Arabind Swain , Sean Alexander Ridout , Ilya Nemenman

We obtain the asymptotic distribution of eigenvalues of real symmetric tridiagonal matrices as their dimension increases to infinity and whose diagonal and off-diagonal elements asymptotically change with the index n as J_{nt+i nt+i}\sim…

Mathematical Physics · Physics 2007-05-23 I. V. Krasovsky

Eigenvalues of Wigner matrices has been a major topic of investigation. A particularly important subclass of such random matrices is formed by the adjacency matrix of an Erd\H{o}s-R\'{e}nyi graph $\mathcal{G}_{n,p}$ equipped with i.i.d.…

Probability · Mathematics 2022-06-15 Shirshendu Ganguly , Ella Hiesmayr , Kyeongsik Nam
‹ Prev 1 2 3 10 Next ›