English
Related papers

Related papers: Universal Approximation of Markov Kernels by Shall…

200 papers

Scalability properties of deep neural networks raise key research questions, particularly as the problems considered become larger and more challenging. This paper expands on the idea of conditional computation introduced by Bengio, et.…

Machine Learning · Computer Science 2014-01-30 Andrew Davis , Itamar Arel

In this paper we study shallow neural network functions which are linear combinations of compositions of activation and quadratic functions, replacing standard affine linear functions, often called neurons. We show the universality of this…

Numerical Analysis · Mathematics 2022-05-12 Leon Frischauf , Otmar Scherzer , Cong Shi

We prove a negative result for the approximation of functions defined on compact subsets of $\mathbb{R}^d$ (where $d \geq 2$) using feedforward neural networks with one hidden layer and arbitrary continuous activation function. In a…

Machine Learning · Computer Science 2020-08-26 J. M. Almira , P. E. Lopez-de-Teruel , D. J. Romero-Lopez , F. Voigtlaender

Single hidden layer feedforward neural networks can represent multivariate functions that are sums of ridge functions. These ridge functions are defined via an activation function and customizable weights. The paper deals with best…

Functional Analysis · Mathematics 2020-11-24 Steffen Goebbels

We study the mixtures of factorizing probability distributions represented as visible marginal distributions in stochastic layered networks. We take the perspective of kernel transitions of distributions, which gives a unified picture of…

Machine Learning · Statistics 2012-11-06 Guido F. Montufar , Jason Morton

This work explores the neural network approximation capabilities for functions within the spectral Barron space $\mathscr{B}^s$, where $s$ is the smoothness index. We demonstrate that for functions in $\mathscr{B}^{1/2}$, a shallow neural…

Numerical Analysis · Mathematics 2025-07-10 Yulei Liao , Pingbing Ming , Hao Yu

We introduce the concept of scalable neural network kernels (SNNKs), the replacements of regular feedforward layers (FFLs), capable of approximating the latter, but with favorable computational properties. SNNKs effectively disentangle the…

Machine Learning · Computer Science 2024-03-07 Arijit Sehanobish , Krzysztof Choromanski , Yunfan Zhao , Avinava Dubey , Valerii Likhosherstov

We present a data-driven method for computing approximate forward reachable sets using separating kernels in a reproducing kernel Hilbert space. We frame the problem as a support estimation problem, and learn a classifier of the support as…

Optimization and Control · Mathematics 2020-11-20 Adam J. Thorpe , Kendric R. Ortiz , Meeko M. K. Oishi

In this paper, we compute finite sample bounds for data-driven approximations of the solution to stochastic reachability problems. Our approach uses a nonparametric technique known as kernel distribution embeddings, and provides…

Optimization and Control · Mathematics 2021-12-09 Adam J. Thorpe , Kendric R. Ortiz , Meeko M. K. Oishi

This paper explores the complexity of deep feedforward networks with linear pre-synaptic couplings and rectified linear activations. This is a contribution to the growing body of work contrasting the representational power of deep and…

Machine Learning · Computer Science 2014-02-17 Razvan Pascanu , Guido Montufar , Yoshua Bengio

Multiplication layers are a key component in various influential neural network modules, including self-attention and hypernetwork layers. In this paper, we investigate the approximation capabilities of deep neural networks with…

Machine Learning · Computer Science 2023-01-12 Ido Ben-Shaul , Tomer Galanti , Shai Dekel

A novel strategy that combines a given collection of $\pi$-reversible Markov kernels is proposed. At each Markov transition, one of the available kernels is selected via a state-dependent probability distribution. In contrast to random-scan…

Methodology · Statistics 2022-03-30 Florian Maire , Pierre Vandekerkhove

This paper develops simple feed-forward neural networks that achieve the universal approximation property for all continuous functions with a fixed finite number of neurons. These neural networks are simple because they are designed with a…

Machine Learning · Computer Science 2022-10-10 Zuowei Shen , Haizhao Yang , Shijun Zhang

George Cybenko's landmark 1989 paper showed that there exists a feedforward neural network, with exactly one hidden layer (and a finite number of neurons), that can arbitrarily approximate a given continuous function $f$ on the unit…

Machine Learning · Computer Science 2019-02-12 Elliott Zaresky-Williams

We prove that any one-dimensional (1D) quantum state with small quantum conditional mutual information in all certain tripartite splits of the system, which we call a quantum approximate Markov chain, can be well-approximated by a Gibbs…

Quantum Physics · Physics 2019-08-13 Kohtaro Kato , Fernando G. S. L. Brandao

The universal approximation theorem asserts that a single hidden layer neural network approximates continuous functions with any desired precision on compact sets. As an existential result, the universal approximation theorem supports the…

Machine Learning · Computer Science 2023-09-15 Wington L. Vital , Guilherme Vieira , Marcos Eduardo Valle

We aim at the construction of a Hidden Markov Model (HMM) of assigned complexity (number of states of the underlying Markov chain) which best approximates, in Kullback-Leibler divergence rate, a given stationary process. We establish, under…

Optimization and Control · Mathematics 2014-07-03 Lorenzo Finesso , Angela Grassi , Peter Spreij

We study feedforward neural networks with inputs from a topological vector space (TVS-FNNs). Unlike traditional feedforward neural networks, TVS-FNNs can process a broader range of inputs, including sequences, matrices, functions and more.…

Machine Learning · Computer Science 2024-09-20 Vugar Ismailov

It is well-known that neural networks are universal approximators, but that deeper networks tend in practice to be more powerful than shallower ones. We shed light on this by proving that the total number of neurons $m$ required to…

Machine Learning · Computer Science 2018-04-30 David Rolnick , Max Tegmark

Hidden Markov Models (HMMs) can be accurately approximated using co-occurrence frequencies of pairs and triples of observations by using a fast spectral method in contrast to the usual slow methods like EM or Gibbs sampling. We provide a…

Machine Learning · Statistics 2012-03-29 Dean P. Foster , Jordan Rodu , Lyle H. Ungar