English
Related papers

Related papers: Improved subsample-and-aggregate via the private m…

200 papers

This paper presents an integrated framework for estimation and inference from generalized linear models using adjusted score equations that result in mean and median bias reduction. The framework unifies theoretical and methodological…

Methodology · Statistics 2019-01-15 Ioannis Kosmidis , Euloge Clovis Kenne Pagui , Nicola Sartori

We consider the problem of communication-constrained collaborative personalized mean estimation under a privacy constraint in an environment of several agents continuously receiving data according to arbitrary unknown agent-specific…

Social and Information Networks · Computer Science 2025-11-10 Yauhen Yakimenka , Hsuan-Yin Lin , Eirik Rosnes , Jörg Kliewer

Many applications, including natural language processing, sensor networks, collaborative filtering, and federated learning, call for estimating discrete distributions from data collected in batches, some of which may be untrustworthy,…

Machine Learning · Computer Science 2020-02-26 Ayush Jain , Alon Orlitsky

We propose a novel Bayesian inference framework for distributed differentially private linear regression. We consider a distributed setting where multiple parties hold parts of the data and share certain summary statistics of their portions…

Machine Learning · Statistics 2023-06-08 Barış Alparslan , Sinan Yıldırım , Ş. İlker Birbil

In situations where the sampling units in a study can be more easily ranked based on the measurement of an auxiliary variable, ranked set sampling provide unbiased estimators for the mean of a population that they are more efficient than…

Statistics Theory · Mathematics 2014-05-13 Saeid Tahmasebi , Ali Akbar Jafari

Rating aggregation plays a crucial role in various fields, such as product recommendations, hotel rankings, and teaching evaluations. However, traditional averaging methods can be affected by participation bias, where some raters do not…

Machine Learning · Computer Science 2025-02-07 Yongkang Guo , Yuqing Kong , Jialiang Liu

Estimating the unknown density from which a given independent sample originates is more difficult than estimating the mean, in the sense that for the best popular non-parametric density estimators, the mean integrated square error converges…

Statistics Theory · Mathematics 2021-09-08 Pierre L'Ecuyer , Florian Puchhammer , Amal Ben Abdellah

In this work, we give efficient algorithms for privately estimating a Gaussian distribution in both pure and approximate differential privacy (DP) models with optimal dependence on the dimension in the sample complexity. In the pure DP…

Data Structures and Algorithms · Computer Science 2023-06-02 Daniel Alabi , Pravesh K. Kothari , Pranay Tankala , Prayaag Venkat , Fred Zhang

Diffusion models now generate high-quality, diverse samples, with an increasing focus on more powerful models. Although ensembling is a well-known way to improve supervised models, its application to unconditional score-based diffusion…

Machine Learning · Computer Science 2026-01-22 Raphaël Razafindralambo , Rémy Sun , Frédéric Precioso , Damien Garreau , Pierre-Alexandre Mattei

Over the years, the most popularly used control chart for statistical process control has been Shewhart's $\bar{X}-S$ or $\bar{X}-R$ chart along with its multivariate generalizations. But, such control charts suffer from the lack of…

Computation · Statistics 2012-11-20 Kushal Kr. Dey , Kumaresh Dhara , Bikram Karmakar , Sukalyan Sengupta

Distributed compressive sensing is a framework considering jointly sparsity within signal ensembles along with multiple measurement vectors (MMVs). The current theoretical bound of performance for MMVs, however, is derived to be the same…

Information Theory · Computer Science 2016-09-12 Sung-Hsien Hsieh , Wei-Jie Liang , Chun-Shien Lu , Soo-Chang Pei

We study the problem of estimating the mean of a random vector $X$ given a sample of $N$ independent, identically distributed points. We introduce a new estimator that achieves a purely sub-Gaussian performance under the only condition that…

Statistics Theory · Mathematics 2017-02-03 Gábor Lugosi , Shahar Mendelson

The problem of estimating the mean of random functions based on discretely sampled data arises naturally in functional data analysis. In this paper, we study optimal estimation of the mean function under both common and independent designs.…

Statistics Theory · Mathematics 2012-02-24 T. Tony Cai , Ming Yuan

Estimation using pooled sampling has long been an area of interest in the group testing literature. Such research has focused primarily on the assumed use of fixed sampling plans (i), although some recent papers have suggested alternative…

Statistics Theory · Mathematics 2017-03-27 Gregory Haber , Yaakov Malinovsky , Paul Albert

A general methodology is introduced for the construction and effective application of control variates to estimation problems involving data from reversible MCMC samplers. We propose the use of a specific class of functions as control…

Computation · Statistics 2010-08-10 Petros Dellaportas , Ioannis Kontoyiannis

When reporting the results of clinical studies, some researchers may choose the five-number summary (including the sample median, the first and third quartiles, and the minimum and maximum values) rather than the sample mean and standard…

Methodology · Statistics 2020-06-18 Jiandong Shi , Dehui Luo , Hong Weng , Xian-Tao Zeng , Lu Lin , Haitao Chu , Tiejun Tong

Importance sampling is a widely used technique to estimate properties of a distribution. This paper investigates trading-off some bias for variance by adaptively winsorizing the importance sampling estimator. The novel winsorizing…

Computation · Statistics 2021-02-10 Paulo Orenstein

In this paper, we observe a sparse mean vector through Gaussian noise and we aim at estimating some additive functional of the mean in the minimax sense. More precisely, we generalize the results of (Collier et al., 2017, 2019) to a very…

Statistics Theory · Mathematics 2019-08-30 Olivier Collier , Laëtitia Comminges

Sparse variable selection improves interpretability and generalization in high-dimensional learning by selecting a small subset of informative features. Recent advances in Mixed Integer Programming (MIP) have enabled solving large-scale…

Machine Learning · Statistics 2025-10-28 Petros Prastakos , Kayhan Behdin , Rahul Mazumder

Randomized (dithered) quantization is a method capable of achieving white reconstruction error independent of the source. Dithered quantizers have traditionally been considered within their natural setting of uniform quantization. In this…

Information Theory · Computer Science 2017-04-26 Emrah Akyol , Kenneth Rose