English
Related papers

Related papers: Large deviation principles for convolutional Bayes…

200 papers

Recent works show an intriguing phenomenon of Frequency Principle (F-Principle) that deep neural networks (DNNs) fit the target function from low to high frequency during the training, which provides insight into the training and…

Machine Learning · Computer Science 2020-10-19 Tao Luo , Zheng Ma , Zhi-Qin John Xu , Yaoyu Zhang

We study an inhomogeneous sparse random graph on [N] = {1, . . . , N } as introduced in a seminal paper by Bollobas, Janson and Riordan (2007): vertices have a type (here in a compact metric space S), and edges between different vertices…

Probability · Mathematics 2023-08-21 Luisa Andreis , Wolfgang König , Heide Langhammer , Robert I. A. Patterson

Despite exceptional predictive performance of Deep sequence models (DSMs), the main concern of their deployment centers around the lack of uncertainty awareness. In contrast, probabilistic models quantify the uncertainty associated with…

Machine Learning · Computer Science 2026-03-03 Wenlong Chen

Generalized Large deviation principles was developed for Colombeau-Ito SDE with a random coefficients. We is significantly expand the classical theory of large deviations for randomly perturbed dynamical systems developed by Freidlin and…

Mathematical Physics · Physics 2024-06-03 Jaykov Foukzon

For an arbitrary negative Schwarzian unimodal map with non-flat critical point, we establish the level-2 Large Deviation Principle (LDP) for empirical distributions. We also give an example of a multimodal map for which the level-2 LDP does…

Dynamical Systems · Mathematics 2026-03-18 Hiroki Takahasi , Masato Tsujii

We introduce a principled approach for unsupervised structure learning of deep neural networks. We propose a new interpretation for depth and inter-layer connectivity where conditional independencies in the input distribution are encoded…

Machine Learning · Statistics 2018-10-18 Raanan Y. Rohekar , Shami Nisimov , Yaniv Gurwicz , Guy Koren , Gal Novik

This article studies the infinite-width limit of deep feedforward neural networks whose weights are dependent, and modelled via a mixture of Gaussian distributions. Each hidden node of the network is assigned a nonnegative random variable…

Machine Learning · Statistics 2025-02-06 Hoil Lee , Fadhel Ayed , Paul Jung , Juho Lee , Hongseok Yang , François Caron

In this paper, we rigorously derive Central Limit Theorems (CLT) for Bayesian two-layerneural networks in the infinite-width limit and trained by variational inference on a regression task. The different networks are trained via different…

Machine Learning · Statistics 2024-06-14 Arnaud Descours , Tom Huix , Arnaud Guillin , Manon Michel , Éric Moulines , Boris Nectoux

Bayesian neural networks (BNNs) have been long considered an ideal, yet unscalable solution for improving the robustness and the predictive uncertainty of deep neural networks. While they could capture more accurately the posterior…

Computer Vision and Pattern Recognition · Computer Science 2021-03-26 Gianni Franchi , Andrei Bursuc , Emanuel Aldea , Severine Dubuisson , Isabelle Bloch

Neural Processes (NPs) are meta-learning models that learn to map sets of observations to approximations of the corresponding posterior predictive distributions. By accommodating variable-sized, unstructured collections of observations and…

Machine Learning · Computer Science 2026-02-10 Peiman Mohseni , Nick Duffield

Following the traditional paradigm of convolutional neural networks (CNNs), modern CNNs manage to keep pace with more recent, for example transformer-based, models by not only increasing model depth and width but also the kernel size. This…

Computer Vision and Pattern Recognition · Computer Science 2023-06-23 Paul Gavrikov , Janis Keuper

We show generalisation error bounds for deep learning with two main improvements over the state of the art. (1) Our bounds have no explicit dependence on the number of classes except for logarithmic factors. This holds even when formulating…

Machine Learning · Computer Science 2021-02-23 Antoine Ledent , Waleed Mustafa , Yunwen Lei , Marius Kloft

This work concerns about multiscale multivalued McKean-Vlasov stochastic systems. First of all, we use a contractive mapping principle to establish the well-posedness for fully coupled multivalued McKean-Vlasov stochastic systems under…

Probability · Mathematics 2025-09-30 Huijie Qiao

We establish a functional large deviation principle for fully connected multi-layer perceptrons with i.i.d. Gaussian weights (LeCun initialization) and general Lipschitz activation functions, including therefore the popular case of ReLU.…

A key property of neural networks driving their success is their ability to learn features from data. Understanding feature learning from a theoretical viewpoint is an emerging field with many open questions. In this work we capture…

Disordered Systems and Neural Networks · Physics 2024-05-20 Kirsten Fischer , Javed Lindner , David Dahmen , Zohar Ringel , Michael Krämer , Moritz Helias

We consider a family of positive operator valued measures associated with representations of compact connected Lie groups. For many independent copies of a single state and a tensor power representation we show that the observed probability…

Mathematical Physics · Physics 2024-09-04 Alonso Botero , Matthias Christandl , Péter Vrana

In this article we obtain large deviation asymptotics for supercritical communication networks modelled as signal-interference-noise ratio networks. To do this, we define the empirical power measure and the empirical connectivity measure,…

Probability · Mathematics 2020-11-13 E. Sakyi-Yeboah , P. S. Andam , L. Asiedu , K. Doku-Amponsah

In decision-making systems, it is important to have classifiers that have calibrated uncertainties, with an optimisation objective that can be used for automated model selection and training. Gaussian processes (GPs) provide uncertainty…

Machine Learning · Statistics 2020-03-05 Vincent Dutordoir , Mark van der Wilk , Artem Artemev , James Hensman

The main results in this paper concern large deviations for families of non-Gaussian processes obtained as suitable perturbations of continuous centered multivariate Gaussian processes which satisfy a large deviation principle. We present…

Probability · Mathematics 2023-07-06 C. Macci , B. Pacchiarotti

Neural network approaches for meta-learning distributions over functions have desirable properties such as increased flexibility and a reduced complexity of inference. Building on the successes of denoising diffusion models for generative…

Machine Learning · Statistics 2023-06-08 Vincent Dutordoir , Alan Saul , Zoubin Ghahramani , Fergus Simpson