English
Related papers

Related papers: Asymptotic Singular Value Distribution of Linear C…

200 papers

Conventional deep convolutional neural networks (CNNs) apply convolution operators uniformly in space across all feature maps for hundreds of layers - this incurs a high computational cost for real-time applications. For many problems such…

Computer Vision and Pattern Recognition · Computer Science 2018-06-08 Mengye Ren , Andrei Pokrovsky , Bin Yang , Raquel Urtasun

We investigate singularly perturbed elliptic problems with multiplicative nonlocal diffusion terms subject to Robin boundary conditions. The diffusion depends on a global quantity of the solution, which introduces a nonlocal coupling…

Analysis of PDEs · Mathematics 2026-04-08 Chiun-Chang Lee , Sang-Hyuck Moon , Wen Yang

An approach to construct explicit integral representations for two-layer ReLU networks is presented, which provides relatively simple representations for any multivariate polynomial. Quantitative bounds are provided for a particular,…

Machine Learning · Statistics 2026-05-13 Anthony Lee

Random Matrix Theory (RMT) has successfully modeled diverse systems, from energy levels of heavy nuclei to zeros of $L$-functions. Many statistics in one can be interpreted in terms of quantities of the other; for example, zeros of…

Singular value decomposition is the key tool in the analysis and understanding of linear regularization methods. In the last decade nonlinear variational approaches such as $\ell^1$ or total variation regularizations became quite prominent…

Numerical Analysis · Mathematics 2012-11-12 Martin Benning , Martin Burger

Singular spectrum analysis is developed as a nonparametric spectral decomposition of a time series. It can be easily extended to the decomposition of multidimensional lattice-like data through the filtering interpretation. In this…

Computer Vision and Pattern Recognition · Computer Science 2015-05-08 Kenji Kume , Naoko Nose-Togawa

Convolutional neural networks (CNNs) have achieved breakthrough performances in a wide range of applications including image classification, semantic segmentation, and object detection. Previous research on characterizing the generalization…

Machine Learning · Statistics 2019-10-04 Shan Lin , Jingwei Zhang

In this paper, we study the problem of approximately computing the product of two real matrices. In particular, we analyze a dimensionality-reduction-based approximation algorithm due to Sarlos [1], introducing the notion of nuclear rank as…

Statistics Theory · Mathematics 2014-04-01 Anastasios Kyrillidis , Michail Vlachos , Anastasios Zouzias

The singular value and spectral distribution of Toeplitz matrix sequences with Lebesgue integrable generating functions is well studied. Early results were provided in the classical Szeg{\H{o}} theorem and the Avram-Parter theorem, in which…

Numerical Analysis · Mathematics 2018-10-08 Sean Hon , Mohammad Ayman Mursaleen , Stefano Serra-Capizzano

Convolutional neural networks are capable of learning powerful representational spaces, which are necessary for tackling complex learning tasks. However, due to the model capacity required to capture such representations, they are often…

Computer Vision and Pattern Recognition · Computer Science 2017-11-30 Terrance DeVries , Graham W. Taylor

Improving information flow in deep networks helps to ease the training difficulties and utilize parameters more efficiently. Here we propose a new convolutional neural network architecture with alternately updated clique (CliqueNet). In…

Computer Vision and Pattern Recognition · Computer Science 2018-04-04 Yibo Yang , Zhisheng Zhong , Tiancheng Shen , Zhouchen Lin

Several regularization methods have recently been introduced which force the latent activations of an autoencoder or deep neural network to conform to either a Gaussian or hyperspherical distribution, or to minimize the implicit rank of the…

Machine Learning · Computer Science 2022-07-04 Xuefeng Li , Alan Blair

Symmetry is present in nature and science. In image processing, kernels for spatial filtering possess some symmetry (e.g. Sobel operators, Gaussian, Laplacian). Convolutional layers in artificial feed-forward neural networks have typically…

Computer Vision and Pattern Recognition · Computer Science 2019-06-12 Gregory Dzhezyan , Hubert Cecotti

Iterative hard thresholding (IHT) has gained in popularity over the past decades in large-scale optimization. However, convergence properties of this method have only been explored recently in non-convex settings. In matrix completion,…

Optimization and Control · Mathematics 2023-01-11 Trung Vu , Evgenia Chunikhina , Raviv Raich

Deep learning models have proven enormously successful at using multiple layers of representation to learn relevant features of structured data. Encoding physical symmetries into these models can improve performance on difficult tasks, and…

Machine Learning · Computer Science 2025-10-21 Cassidy Ashworth , Pietro Liò , Francesco Caso

There has been a recent surge of interest in the study of asymptotic reconstruction performance in various cases of generalized linear estimation problems in the teacher-student setting, especially for the case of i.i.d standard normal…

Machine Learning · Statistics 2023-02-20 Cedric Gerbelot , Alia Abbara , Florent Krzakala

We propose a novel image sampling method for differentiable image transformation in deep neural networks. The sampling schemes currently used in deep learning, such as Spatial Transformer Networks, rely on bilinear interpolation, which…

Computer Vision and Pattern Recognition · Computer Science 2019-09-11 Wei Jiang , Weiwei Sun , Andrea Tagliasacchi , Eduard Trulls , Kwang Moo Yi

In this note, we are concerned with the asymptotic approximation of a class of double integrals which can be represented as an angular spectrum superposition. These double integrals typically appear in electromagnetic scattering problems.…

Mathematical Physics · Physics 2007-05-23 Fei Wang

Typical convolutional neural networks (CNNs) have several millions of parameters and require a large amount of annotated data to train them. In medical applications where training data is hard to come by, these sophisticated machine…

Computer Vision and Pattern Recognition · Computer Science 2016-12-09 Rahul Venkataramani , Sheshadri Thiruvenkadam , Prasad Sudhakar , Hariharan Ravishankar , Vivek Vaidya

A butterfly network consists of logarithmically many layers, each with a linear number of non-zero weights (pre-specified). The fast Johnson-Lindenstrauss transform (FJLT) can be represented as a butterfly network followed by a projection…

Machine Learning · Computer Science 2021-07-06 Nir Ailon , Omer Leibovich , Vineet Nair