Related papers: On some discrete random variables arising from rec…
Advances in neural variational inference have facilitated the learning of powerful directed graphical models with continuous latent variables, such as variational autoencoders. The hope is that such models will learn to represent rich,…
For any $n > 0$ and $0 \leq m < n$, let $P_{n,m}$ be the poset of projective equivalence classes of $\{-,0,+\}$-vectors of length $n$ with sign variation bounded by $m$, ordered by reverse inclusion of the positions of zeros. Let…
Given a sequence $(M_{k}, Q_{k})_{k\ge 1}$ of independent, identically distributed ran\-dom vectors with nonnegative components, we consider the recursive Markov chain $(X_{n})_{n\ge 0}$, defined by the random difference equation…
Let $B(n,p)$ denote a binomial random variable with parameters $n$ and $p$. Vas\v{e}k Chv\'{a}tal conjectured that for any fixed $n\geq 2$, as $m$ ranges over $\{0,\ldots,n\}$, the probability $q_m:=P(B(n,m/n)\leq m)$ is the smallest when…
Let $X, X_1, X_2,...$ be a sequence of non-degenerate i.i.d. random variables with mean zero. The best possible weighted approximations are investigated in $D[0, 1]$ for the partial sum processes $\{S_{[nt]}, 0\le t\le 1\}$, where…
Observed associations in a database may be due in whole or part to variations in unrecorded (latent) variables. Identifying such variables and their causal relationships with one another is a principal goal in many scientific and practical…
For a real-valued sequence $(x_n)_{n=1}^\infty$, denote by $S_N(\ell)$ the number of its first $N$ fractional parts lying in a random interval of size $\ell:=L/N$, where $L=o(N)$ as $N\to\infty$. We study the variance of $S_N(\ell)$ (the…
We consider the problem of testing whether pairs of univariate random variables are associated. Few tests of independence exist that are consistent against all dependent alternatives and are distribution free. We propose novel tests that…
We investigate one/two-sample mean tests for high-dimensional compositional data when the number of variables is comparable with the sample size, as commonly encountered in microbiome research. Existing methods mainly focus on max-type test…
The paper is devoted to infinite Bernoulli convolutions generated by positive multigeometric series and to probability distributions of random variables whose digits in an even integer base-$s$ expansion with two redundant digits form a…
Motivation: Algorithms that discover variables which are causally related to a target may inform the design of experiments. With observational gene expression data, many methods discover causal variables by measuring each variable's degree…
We proved that for any finite collection of sparse subgraphs $(D_m)_{m=1}^\ell$ of the complete graph $K_{2n}$, and a uniformly chosen perfect matching $R$ in $K_{2n}$, the random vector $(|E(R \cap D_m)|)_{m=1}^\ell$ jointly converges to a…
We consider random vectors $X$ that satisfy the equation in law $X=AX+B$, where $A$ is a given random diagonal matrix and $B$ a given random vector, both independent of $X$. It is well known by the works of Kesten and Goldie that the…
The problem of statistical learning is to construct a predictor of a random variable $Y$ as a function of a related random variable $X$ on the basis of an i.i.d. training sample from the joint distribution of $(X,Y)$. Allowable predictors…
Constant (naive) imputation is still widely used in practice as this is a first easy-to-use technique to deal with missing data. Yet, this simple method could be expected to induce a large bias for prediction purposes, as the imputed input…
The problem is sequence prediction in the following setting. A sequence x1,..., xn,... of discrete-valued observations is generated according to some unknown probabilistic law (measure) mu. After observing each outcome, it is required to…
We obtain some new results concerning the small deviation problem for $S=\sum_n q^n X_n$ and $M=\sup_n q^n X_n$, where $0<q<1$ and $(X_n)$ are i.i.d. non-negative random variables. In particular, the asymptotics is shown to be the same for…
We study asymptotic behavior of the moments $M_k(\lambda)$ of the sum $X_1+\dots+X_{N_\lambda}$, where $N_\lambda$ follows the Poisson probability distribution with mean value $\lambda$ and $\{X_j\}$ is a family of i.i.d. random variables…
Improving Importance Sampling estimators for rare event probabilities requires sharp approximations of conditional densities. This is achieved for events E_{n}:=(f(X_{1})+...+f(X_{n}))\inA_{n} where the summands are i.i.d. and E_{n} is a…
We prove the following type of discrete entropy monotonicity for sums of isotropic, log-concave, independent and identically distributed random vectors $X_1,\dots,X_{n+1}$ on $\mathbb{Z}^d$: $$ H(X_1+\cdots+X_{n+1}) \geq H(X_1+\cdots+X_{n})…