English
Related papers

Related papers: Universal Approximation of Markov Kernels by Shall…

200 papers

We report an exact likelihood computation for Linear Gaussian Markov processes that is more scalable than existing algorithms for complex models and sparsely sampled signals. Better scaling is achieved through elimination of repeated…

Machine Learning · Statistics 2018-05-21 Stijn de Waele

We study the approximation of a Markov chain on a reduced state space, for both discrete- and continuous-time Markov chains. In this context, we extend the existing theory of formal error bounds for the approximated transient distributions.…

Probability · Mathematics 2025-02-12 Fabian Michel , Markus Siegle

We propose an adaptive estimator for the stationary distribution of a bifurcating Markov Chain on $\mathbb R^d$. Bifurcating Markov chains (BMC for short) are a class of stochastic processes indexed by regular binary trees. A kernel…

Statistics Theory · Mathematics 2017-06-22 S Valere Bitseki Penda , Angelina Roche

A topological neural network (TNN), which takes data from a Tychonoff topological space instead of the usual finite dimensional space, is introduced. As a consequence, a distributional neural network (DNN) that takes Borel measures as data…

Machine Learning · Computer Science 2023-05-29 Michael A. Kouritzin , Daniel Richard

We propose o1Neuro, a new neural network model built on sparse indicator activation neurons, with two key statistical properties. (1) Constructive universal approximation: At the population level, a deep o1Neuro can approximate any…

Machine Learning · Statistics 2025-09-15 Chien-Ming Chi

This note provides a family of classification problems, indexed by a positive integer $k$, where all shallow networks with fewer than exponentially (in $k$) many nodes exhibit error at least $1/6$, whereas a deep network with 2 nodes in…

Machine Learning · Computer Science 2015-09-30 Matus Telgarsky

We present an efficient exact algorithm for estimating state sequences from outputs (or observations) in imprecise hidden Markov models (iHMM), where both the uncertainty linking one state to the next, and that linking a state to its…

Artificial Intelligence · Computer Science 2012-10-08 Jasper De Bock , Gert de Cooman

We study the following two related problems. The first is to determine to what error an arbitrary zonoid in $\mathbb{R}^{d+1}$ can be approximated in the Hausdorff distance by a sum of $n$ line segments. The second is to determine optimal…

Machine Learning · Statistics 2025-03-25 Jonathan W. Siegel

Composite adaptive radial basis function neural network (RBFNN) control with a lattice distribution of hidden nodes has three inherent demerits: 1) the approximation domain of adaptive RBFNNs is difficult to be determined a priori; 2) only…

Systems and Control · Electrical Eng. & Systems 2021-04-23 Qiong Liu , Dongyu Li , Shuzhi Sam Ge , Zhong Ouyang

We study additive mixtures of Markov kernels of the form $A_\alpha = \alpha P + (1-\alpha)G$, where $\alpha \in [0,1]$, $P$ is a baseline sampler and $G$ is a Gibbs kernel induced by a partition of the state space. We first motivate the…

Probability · Mathematics 2026-04-15 Ryan J. Y. Lim , Michael C. H. Choi

Nonparametric identification and maximum likelihood estimation for finite-state hidden Markov models are investigated. We obtain identification of the parameters as well as the order of the Markov chain if the transition probability…

Statistics Theory · Mathematics 2015-10-01 Grigory Alexandrovich , Hajo Holzmann , Anna Leister

In this paper, we study the sample complexity lower bounds for the exact recovery of parameters and for a positive excess risk of a feed-forward, fully-connected neural network for binary classification, using information-theoretic tools.…

Machine Learning · Statistics 2020-10-30 Xiaochen Yang , Jean Honorio

Training deep neural networks with the error backpropagation algorithm is considered implausible from a biological perspective. Numerous recent publications suggest elaborate models for biologically plausible variants of deep learning,…

Neural and Evolutionary Computing · Computer Science 2019-07-16 Bernd Illing , Wulfram Gerstner , Johanni Brea

We investigate the problem of quantifying contraction coefficients of Markov transition kernels in Kantorovich ($L^1$ Wasserstein) distances. For diffusion processes, relatively precise quantitative bounds on contraction rates have recently…

Probability · Mathematics 2018-08-22 Andreas Eberle , Mateusz B. Majka

This paper presents two main theoretical results concerning shallow neural networks with ReLU$^k$ activation functions. We establish a novel integral representation for Sobolev spaces, showing that every function in…

Numerical Analysis · Mathematics 2025-05-13 Xinliang Liu , Tong Mao , Jinchao Xu

Markov state models (MSMs) have been successful in computing metastable states, slow relaxation timescales and associated structural changes, and stationary or kinetic experimental observables of complex molecules from large amounts of…

Chemical Physics · Physics 2015-06-17 Frank Noe , Hao Wu , Jan-Hendrik Prinz , Nuria Plattner

This study explores the number of neurons required for a Rectified Linear Unit (ReLU) neural network to approximate multivariate monomials. We establish an exponential lower bound on the complexity of any shallow network approximating the…

Machine Learning · Computer Science 2023-05-17 Itai Shapira

We derive explicit upper bounds for the $\bar{d}$-distance between a chain of infinite order and its canonical $k$-steps Markov approximation. Our proof is entirely constructive and involves a "coupling from the past" argument. The new…

Probability · Mathematics 2012-01-16 Sandro Gallo , Matthieu Lerasle , Daniel Yasumasa Takahashi

Consider the problem of predicting the next symbol given a sample path of length n, whose joint distribution belongs to a distribution class that may have long-term memory. The goal is to compete with the conditional predictor that knows…

Statistics Theory · Mathematics 2024-04-25 Yanjun Han , Tianze Jiang , Yihong Wu

We introduce a distance between kernels based on the Wasserstein distances between their values, study its properties, and prove that it is a metric on an appropriately defined space of kernels. We also relate it to various modes of…

Optimization and Control · Mathematics 2024-01-29 Zhengqi Lin , Andrzej Ruszczynski