English
Related papers

Related papers: Concentration inequalities for the sum in sampling…

200 papers

Let $X$ be a random variable with distribution function $F,$ and $X_{1},X_{2},...,X_{n}$ are independent copies of $X.$ Consider the order statistics $X_{i:n},$ $i=1,2,...,n$ and denote $F_{i:n}(x)=P\{X_{i:n}\leq x\}.$ Using majorization…

Statistics Theory · Mathematics 2011-09-02 Ismihan Bairamov

This paper studies one-sided hypothesis testing under random sampling without replacement. That is, when $n+1$ binary random variables $X_1,\ldots, X_{n+1}$ are subject to a permutation invariant distribution and $n$ binary random variables…

Statistics Theory · Mathematics 2022-11-07 Zihao Li , Huangjun Zhu , Masahito Hayashi

Concentration inequalities quantify the deviation of a random variable from a fixed value. In spite of numerous applications, such as opinion surveys or ecological counting procedures, few concentration results are known for the setting of…

Statistics Theory · Mathematics 2015-07-28 Rémi Bardenet , Odalric-Ambrym Maillard

The phenomenon of entropy concentration provides strong support for the maximum entropy method, MaxEnt, for inferring a probability vector from information in the form of constraints. Here we extend this phenomenon, in a discrete setting,…

Information Theory · Computer Science 2021-01-11 Kostas N. Oikonomou

We introduce a new generalization of relative entropy to non-negative vectors with sums $\gt 1$. We show in a purely combinatorial setting, with no probabilistic considerations, that in the presence of linear constraints defining a convex…

Information Theory · Computer Science 2024-05-08 Kostas N. Oikonomou

The average properties of the well-known Subset Sum Problem can be studied by the means of its randomised version, where we are given a target value $z$, random variables $X_1, \ldots, X_n$, and an error parameter $\varepsilon > 0$, and we…

For noncorrelated random variables, we study a concentration property of the family of distributions of normalized sums formed by sequences of times of a given large length.

Probability · Mathematics 2007-05-23 Sergey G. Bobkov

Starting with a set of weighted items, we want to create a generic sample of a certain size that we can later use to estimate the total weight of arbitrary subsets. For this purpose, we propose priority sampling which tested on Internet…

Data Structures and Algorithms · Computer Science 2007-05-23 Nick Duffield , Carsten Lund , Mikkel Thorup

Let $S$ be a finite set, and $X_1,\ldots,X_n$ an i.i.d. uniform sample from $S$. To estimate the size $|S|$, without further structure, one can wait for repeats and use the birthday problem. This requires a sample size of the order…

Statistics Theory · Mathematics 2026-04-28 Sourav Chatterjee , Persi Diaconis , Susan Holmes

We study certain sequences involving sums of powers of positive integers and in connection with this, we give examples to show that power majorization does not imply majorization.

Classical Analysis and ODEs · Mathematics 2015-06-26 Peng Gao

In this paper we deal with the problem of testing for the quality of $k$ probability distributions. We introduce a generalization of the maximum mean discrepancy that permits to characterize the null hypothesis. Then, an estimator of it is…

Statistics Theory · Mathematics 2018-11-26 Armando Sosthene Kali Balogoun , Guy Martial Nkiet , Carlos Ogouyandjou

This paper presents a novel algorithm solving the classic problem of generating a random sample of size s from population of size n with non-uniform probabilities. The sampling is done with replacement. The algorithm requires constant…

Data Structures and Algorithms · Computer Science 2016-11-03 Michał Startek

The method of maximum entropy is quite a powerful tool to solve the generalized moment problem, which consists of determining the probability density of a random variable X from the knowledge of the expected values of a few functions of the…

Statistics Theory · Mathematics 2015-10-15 Henryk Gzyl

Let ($X,Y)$ be a random vector with distribution function $F(x,y),$ and $(X_{1},Y_{1}),(X_{2},Y_{2}),...,(X_{n},Y_{n})$ are independent copies of ($X,Y).$ Let $X_{i:n}$ be the $i$th order statistics constructed from the sample…

Statistics Theory · Mathematics 2011-09-08 Ismihan Bairamov

We develop large sample theory for merged data from multiple sources. Main statistical issues treated in this paper are (1) the same unit potentially appears in multiple datasets from overlapping data sources, (2) duplicated items are not…

Statistics Theory · Mathematics 2018-05-22 Takumi Saegusa

The Central Limit Theorem provides a foundation for inferential statistics and hypothesis testing. It describes how standardized statistics behave under repeated sampling from large populations. However, if the size of the sample (n)…

Methodology · Statistics 2026-05-19 Mike Crowhurst

The aim of this paper is twofold. First, three theoretical principles are formalized: randomization, overrepresentation and restriction. We develop these principles and give a rationale for their use in choosing the sampling design in a…

Methodology · Statistics 2016-12-16 Yves Tillé , Matthieu Wilhelm

We prove a lower estimate on the increase in entropy when two copies of a conditional random variable $X | Y$, with $X$ supported on $\mathbb{Z}_q=\{0,1,\dots,q-1\}$ for prime $q$, are summed modulo $q$. Specifically, given two i.i.d copies…

Information Theory · Computer Science 2014-11-27 Venkatesan Guruswami , Ameya Velingker

This paper studies the sample complexity of searching over multiple populations. We consider a large number of populations, each corresponding to either distribution P0 or P1. The goal of the search problem studied here is to find one…

Information Theory · Computer Science 2016-11-17 Matthew L. Malloy , Gongguo Tang , Robert D. Nowak

We investigate the approximation for computing the sum $a_1+...+a_n$ with an input of a list of nonnegative elements $a_1,..., a_n$. If all elements are in the range $[0,1]$, there is a randomized algorithm that can compute an…

Data Structures and Algorithms · Computer Science 2012-03-01 Bin Fu , Wenfeng Li , Zhiyong Peng
‹ Prev 1 2 3 10 Next ›