English
Related papers

Related papers: Learning Sums of Independent Random Variables with…

200 papers

Given a large set $U$ where each item $a\in U$ has weight $w(a)$, we want to estimate the total weight $W=\sum_{a\in U} w(a)$ to within factor of $1\pm\varepsilon$ with some constant probability $>1/2$. Since $n=|U|$ is large, we want to do…

Data Structures and Algorithms · Computer Science 2021-10-29 Lorenzo Beretta , Jakub Tětek

In this paper, we obtain general representations for the joint distributions and copulas of arbitrary dependent random variables absolutely continuous with respect to the product of given one-dimensional marginal distributions. The…

Statistics Theory · Mathematics 2016-08-16 Victor H. de la Peña , Rustam Ibragimov , Shaturgun Sharakhmetov

This work concentrates on the study of inverse determinant sums, which arise from the union bound on the error probability, as a tool for designing and analyzing algebraic space-time block codes. A general framework to study these sums is…

Information Theory · Computer Science 2013-10-01 Roope Vehkalahti , Hsiao-feng Lu , Laura Luzzi

The independence clustering problem is considered in the following formulation: given a set $S$ of random variables, it is required to find the finest partitioning $\{U_1,\dots,U_k\}$ of $S$ into clusters such that the clusters…

Machine Learning · Computer Science 2017-03-21 Daniil Ryabko

In this paper, we prove a conditional limit theorem for independent not necessarily identically distributed random variables. Namely, we obtain the asymptotic distribution of a large number of them given the sum.

Statistics Theory · Mathematics 2020-11-12 Dimbihery Rabenoro

We consider the problem of learning the dynamics of a linear system when one has access to data generated by an auxiliary system that shares similar (but not identical) dynamics, in addition to data from the true system. We use a weighted…

Machine Learning · Statistics 2025-01-20 Lei Xin , Lintao Ye , George Chiu , Shreyas Sundaram

We present randomized algorithms to compute the sumset (Minkowski sum) of two integer sets, and to multiply two univariate integer polynomials given by sparse representations. Our algorithm for sumset has cost softly linear in the combined…

Symbolic Computation · Computer Science 2015-04-27 Andrew Arnold , Daniel S. Roche

For high volume data streams and large data warehouses, sampling is used for efficient approximate answers to aggregate queries over selected subsets. Mathematically, we are dealing with a set of weighted items and want to support queries…

Data Structures and Algorithms · Computer Science 2007-05-23 Mario Szegedy , Mikkel Thorup

In this paper, the joint distribution of the sum and maximum of independent, not necessarily identically distributed, nonnegative random variables is studied for two cases: i) continuous and ii) discrete random variables. First, a recursive…

Probability · Mathematics 2024-07-01 Christos N. Efrem

In most papers establishing consistency for learning algorithms it is assumed that the observations used for training are realizations of an i.i.d. process. In this paper we go far beyond this classical framework by showing that support…

Machine Learning · Statistics 2007-07-04 Ingo Steinwart , Don Hush , Clint Scovel

A popular approach for testing if two univariate random variables are statistically independent consists of partitioning the sample space into bins, and evaluating a test statistic on the binned data. The partition size matters, and the…

Methodology · Statistics 2016-04-28 Ruth Heller , Yair Heller , Shachar Kaufman , Barak Brill , Malka Gorfine

We consider a basic problem in unsupervised learning: learning an unknown \emph{Poisson Binomial Distribution}. A Poisson Binomial Distribution (PBD) over $\{0,1,\dots,n\}$ is the distribution of a sum of $n$ independent Bernoulli random…

Data Structures and Algorithms · Computer Science 2015-02-18 Constantinos Daskalakis , Ilias Diakonikolas , Rocco A. Servedio

We provide rigorous guarantees on learning with the weighted trace-norm under arbitrary sampling distributions. We show that the standard weighted trace-norm might fail when the sampling distribution is not a product distribution (i.e. when…

Machine Learning · Computer Science 2011-06-23 Rina Foygel , Ruslan Salakhutdinov , Ohad Shamir , Nathan Srebro

In this paper, a sparsity-aware adaptive algorithm for distributed learning in diffusion networks is developed. The algorithm follows the set-theoretic estimation rationale. At each time instance and at each node of the network, a closed…

Information Theory · Computer Science 2015-06-03 Symeon Chouvardas , Konstantinos Slavakis , Yannis Kopsinis , Sergios Theodoridis

To learn (statistical) dependencies among random variables requires exponentially large sample size in the number of observed random variables if any arbitrary joint probability distribution can occur. We consider the case that sparse data…

Machine Learning · Computer Science 2007-05-23 Dominik Janzing , Daniel Herrmann

We give a simple inequality for the sum of independent bounded random variables. This inequality improves on the celebrated result of Hoeffding in a special case. It is optimal in the limit where the sum tends to a Poisson random variable.

Probability · Mathematics 2012-10-25 Christopher R. Dance

We generalize the "indirect learning" technique of Furst et. al., 1991 to reduce from learning a concept class over a samplable distribution $\mu$ to learning the same concept class over the uniform distribution. The reduction succeeds when…

Machine Learning · Computer Science 2021-12-24 Eric Binnendyk , Marco Carmosino , Antonina Kolokolova , Ramyaa Ramyaa , Manuel Sabin

Compressive learning forms the exciting intersection between compressed sensing and statistical learning where one exploits forms of sparsity and structure to reduce the memory and/or computational complexity of the learning task. In this…

Machine Learning · Statistics 2021-10-18 Michael P. Sheehan , Mike E. Davies

Comparing the differences in outcomes (that is, in "dependent variables") between two subpopulations is often most informative when comparing outcomes only for individuals from the subpopulations who are similar according to "independent…

Methodology · Statistics 2021-12-20 Mark Tygert

This paper considers how to measure the magnitude of the sum of independent random variables in several ways. We give a formula for the tail distribution for sequences that satisfy the so called Levy property. We then give a connection…

Probability · Mathematics 2007-05-23 Pawel Hitczenko , Stephen Montgomery-Smith