English
Related papers

Related papers: Large deviation principles for convolutional Bayes…

200 papers

There is a previously identified equivalence between wide fully connected neural networks (FCNs) and Gaussian processes (GPs). This equivalence enables, for instance, test set predictions that would have resulted from a fully Bayesian,…

We establish a large deviation principle (LDP) for probability graphons, which are symmetric functions from the unit square into the space of probability measures. This notion extends classical graphons and provides a flexible framework for…

Probability · Mathematics 2025-09-18 Pierfrancesco Dionigi , Giulio Zucal

Bayesian inference is known to provide a general framework for incorporating prior knowledge or specific properties into machine learning models via carefully choosing a prior distribution. In this work, we propose a new type of prior…

Machine Learning · Statistics 2019-02-20 Andrei Atanov , Arsenii Ashukha , Kirill Struminsky , Dmitry Vetrov , Max Welling

Localized sufficient conditions for the large deviation principle of the given stochastic differential equations will be presented for stochastic differential equations with non-Lipschitzian and time-inhomogeneous coefficients, which is…

Probability · Mathematics 2014-04-08 Yunjiao Hu , Guangqiang Lan

We show that the output of a (residual) convolutional neural network (CNN) with an appropriate prior over the weights and biases is a Gaussian process (GP) in the limit of infinitely many convolutional filters, extending similar results for…

Machine Learning · Statistics 2019-05-07 Adrià Garriga-Alonso , Carl Edward Rasmussen , Laurence Aitchison

We continue the development, started in of the asymptotic description of certain stochastic neural networks. We use the Large Deviation Principle (LDP) and the good rate function H announced there to prove that H has a unique minimum mu_e,…

Probability · Mathematics 2014-07-10 Olivier Faugeras , James MacLaurin

We give a Large Deviation Principle (LDP) with explicit rate function for the distribution of vertex degrees in plane trees, a combinatorial model of RNA secondary structures. We calculate the typical degree distributions based on nearest…

Biomolecules · Quantitative Biology 2008-03-28 Yuri Bakhtin , Christine E. Heitsch

The aim of the paper is to establish a large deviation principle (LDP) for the empirical measure of mean-field interacting diffusions in a random environment. The point is to derive such a result once the environment has been frozen…

Probability · Mathematics 2017-03-08 Eric Luçon

We consider fully connected and feedforward deep neural networks with dependent and possibly heavy-tailed weights, as introduced in [26], to address limitations of the standard Gaussian prior. It has been proved in [26] that, as the number…

Machine Learning · Statistics 2026-05-14 Nicola Apollonio , Giovanni Franzina , Giovanni Luca Torrisi

In this paper we derive a Large Deviation Principle (LDP) for inhomogeneous U/V-statistics of a general order. Using this, we derive a LDP for two types of statistics: random multilinear forms, and number of monochromatic copies of a…

Probability · Mathematics 2026-04-01 Sohom Bhattacharya , Nabarun Deb , Sumit Mukherjee

It has long been known that a single-layer fully-connected neural network with an i.i.d. prior over its parameters is equivalent to a Gaussian process (GP), in the limit of infinite network width. This correspondence enables exact Bayesian…

We establish large deviation principle (LDP) for the family of vector-valued random processes $(X^\epsilon,Y^\epsilon),\epsilon\to 0$ defined as $$ X^\epsilon_t=\frac{1}{\epsilon^\kappa}\int_0^t H(\xi^\epsilon_s,Y^\epsilon_s)ds,…

Probability · Mathematics 2016-09-07 A. Guillin , R. Liptser

We study the distributional properties of linear neural networks with random parameters in the context of large networks, where the number of layers diverges in proportion to the number of neurons per layer. Prior works have shown that in…

Machine Learning · Statistics 2024-11-26 Federico Bassetti , Lucia Ladelli , Pietro Rotondo

In this paper we introduce a new notion of convergence of sparse graphs which we call Large Deviations or LD-convergence and which is based on the theory of large deviations. The notion is introduced by "decorating" the nodes of the graph…

Probability · Mathematics 2013-02-20 Christian Borgs , Jennifer Chayes , David Gamarnik

The Large Deviations Principle (LDP) is verified for a homogeneous diffusion process with respect to a Brownian motion $B_t$, $$ X^\eps_t=x_0+\int_0^tb(X^\eps_s)ds+ \eps\int_0^t\sigma(X^\eps_s)dB_s, $$ where $b(x)$ and $\sigma(x)$ are are…

Probability · Mathematics 2011-08-24 P. Chigansky , R. Liptser

We consider temporal models of rapidly changing Markovian networks modulated by time-evolving spatially dependent kernels that define rates for edge formation and dissolution. Alternatively, these can be viewed as Markovian networks with…

Probability · Mathematics 2025-06-11 Shankar Bhamidi , Amarjit Budhiraja , Souvik Ray

Currently there exists rather promising new trend in machine leaning (ML) based on the relationship between neural networks (NN) and Gaussian processes (GP), including many related subtopics, e.g., signal propagation in NNs, theoretical…

Machine Learning · Computer Science 2023-03-01 Andrey Demichev , Alexander Kryukov

Convolutional neural networks (CNNs) have achieved remarkable performance in various fields, particularly in the domain of computer vision. However, why this architecture works well remains to be a mystery. In this work we move a small step…

Machine Learning · Computer Science 2019-05-27 Bing Yu , Junzhao Zhang , Zhanxing Zhu

The Large Deviation Principle is established for stochastic models defined by past-dependent non linear recursions with small noise. In the Markov case we use the result to obtain an explicit expression for the asymptotics of exit time.

Probability · Mathematics 2007-05-23 F. Klebaner , R. Liptser

Inspired by the success of Convolutional Neural Networks (CNNs) for supervised prediction in images, we design the Deconvolutional Generative Model (DGM), a new probabilistic generative model whose inference calculations correspond to those…

Computer Vision and Pattern Recognition · Computer Science 2019-12-10 Tan Nguyen , Nhat Ho , Ankit Patel , Anima Anandkumar , Michael I. Jordan , Richard G. Baraniuk