English
Related papers

Related papers: Large deviation principles for convolutional Bayes…

200 papers

Despite the success of convolutional neural networks (CNNs) in numerous computer vision tasks and their extraordinary generalization performances, several attempts to predict the generalization errors of CNNs have only been limited to a…

Computer Vision and Pattern Recognition · Computer Science 2025-07-17 Vamshi C. Madala , Shivkumar Chandrasekaran , Jason Bunk

The event of large losses plays an important role in credit risk. As these large losses are typically rare, and portfolios usually consist of a large number of positions, large deviation theory is the natural tool to analyze the tail…

Probability · Mathematics 2014-07-03 Vincent Leijdekker , Michel Mandjes , Peter Spreij

Convolutional dictionary learning (CDL) estimates shift invariant basis adapted to multidimensional data. CDL has proven useful for image denoising or inpainting, as well as for pattern discovery on multivariate signals. As estimated…

Machine Learning · Computer Science 2019-01-29 Thomas Moreau , Alexandre Gramfort

Generalization is essential for deep learning. In contrast to previous works claiming that Deep Neural Networks (DNNs) have an implicit regularization implemented by the stochastic gradient descent, we demonstrate explicitly Bayesian…

Machine Learning · Computer Science 2019-10-23 Xinjie Lan , Kenneth E. Barner

We establish a sharp large deviation principle for renewal-reward processes, supposing that each renewal involves a broad-sense reward taking values in a real separable Banach space. In fact, we demonstrate a weak large deviation principle…

Probability · Mathematics 2023-04-24 Marco Zamparo

PCANet and its variants provided good accuracy results for classification tasks. However, despite the importance of network depth in achieving good classification accuracy, these networks were trained with a maximum of nine layers. In this…

Computer Vision and Pattern Recognition · Computer Science 2023-01-30 Mubarakah Alotaibi , Richard Wilson

Continuous input signals like images and time series that are irregularly sampled or have missing values are challenging for existing deep learning methods. Coherently defined feature representations must depend on the values in unobserved…

Machine Learning · Computer Science 2020-10-22 Marc Finzi , Roberto Bondesan , Max Welling

The scaling limit where both the size of the training set $P$ and the width $N$ of a deep neural network grow at the same rate, the so-called proportional-width regime, has been intensely studied for shallow, single-hidden-layer networks.…

This paper presents a novel approach combining convolutional layers (CLs) and large-margin metric learning for training supervised models on small datasets for texture classification. The core of such an approach is a loss function that…

Computer Vision and Pattern Recognition · Computer Science 2022-06-20 Jonathan de Matos , Luiz Eduardo Soares de Oliveira , Alceu de Souza Britto Junior , Alessandro Lameiras Koerich

Let $\sigma(u)$, $u\in \mathbb{R}$ be an ergodic stationary Markov chain, taking a finite number of values $a_1,...,a_m$, and $b(u)=g(\sigma(u))$, where $g$ is a bounded and measurable function. We consider the diffusion type process $$…

Probability · Mathematics 2011-08-24 P. Chigansky , R. Liptser

The theory of stochastic approximations form the theoretical foundation for studying convergence properties of many popular recursive learning algorithms in statistics, machine learning and statistical physics. Large deviations for…

Probability · Mathematics 2025-02-05 Henrik Hult , Adam Lindhe , Pierre Nyquist , Guo-Jhen Wu

Generalization theory has been established for sparse deep neural networks under high-dimensional regime. Beyond generalization, parameter estimation is also important since it is crucial for variable selection and interpretability of deep…

Machine Learning · Statistics 2024-06-27 Dongya Wu , Xin Li

Why heavily parameterized neural networks (NNs) do not overfit the data is an important long standing open question. We propose a phenomenological model of the NN training to explain this non-overfitting puzzle. Our linear frequency…

Machine Learning · Computer Science 2021-05-26 Yaoyu Zhang , Tao Luo , Zheng Ma , Zhi-Qin John Xu

The dominating NLP paradigm of training a strong neural predictor to perform one task on a specific dataset has led to state-of-the-art performance in a variety of applications (eg. sentiment classification, span-prediction based question…

Computation and Language · Computer Science 2021-09-06 Paul Michel

Let $p\in[1,\infty]$. Consider the projection of a uniform random vector from a suitably normalized $\ell^p$ ball in $\mathbb{R}^n$ onto an independent random vector from the unit sphere. We show that sequences of such random projections,…

Probability · Mathematics 2015-12-17 Nina Gantert , Steven Soojin Kim , Kavita Ramanan

A learning-based posterior distribution estimation method, Probabilistic Dipole Inversion (PDI), is proposed to solve quantitative susceptibility mapping (QSM) inverse problem in MRI with uncertainty estimation. A deep convolutional neural…

Image and Video Processing · Electrical Eng. & Systems 2020-04-28 Jinwei Zhang , Hang Zhang , Mert Sabuncu , Pascal Spincemaille , Thanh Nguyen , Yi Wang

Large batch size training in deep neural networks (DNNs) possesses a well-known 'generalization gap' that remarkably induces generalization performance degradation. However, it remains unclear how varying batch size affects the structure of…

Machine Learning · Computer Science 2020-12-17 Fengli Gao , Huicai Zhong

A number of backpropagation-based approaches such as DeConvNets, vanilla Gradient Visualization and Guided Backpropagation have been proposed to better understand individual decisions of deep convolutional neural networks. The saliency maps…

Computer Vision and Pattern Recognition · Computer Science 2019-09-04 Jindong Gu , Yinchong Yang , Volker Tresp

In this work, we investigate the value of employing deep learning for the task of wireless signal modulation recognition. Recently in [1], a framework has been introduced by generating a dataset using GNU radio that mimics the imperfections…

Machine Learning · Computer Science 2018-01-08 Xiaoyu Liu , Diyu Yang , Aly El Gamal

The article obtains large deviation asymptotic for sub-critical communication networks modelled as signal-interference-noise-ratio(SINR) random networks. To achieve this, we define the empirical power measure and the empirical connectivity…

Probability · Mathematics 2021-04-14 E. Sakyi-Yeboah , P. S. Andam , L. Asiedu , K. Doku-Amponsah