English
Related papers

Related papers: Efficient Statistics, in High Dimensions, from Tru…

200 papers

Completely random measures provide a principled approach to creating flexible unsupervised models, where the number of latent features is infinite and the number of features that influence the data grows with the size of the data set. Due…

Machine Learning · Statistics 2020-06-26 Peiyuan Zhu , Alexandre Bouchard-Côté , Trevor Campbell

We give the first polynomial-time algorithm to estimate the mean of a $d$-variate probability distribution with bounded covariance from $\tilde{O}(d)$ independent samples subject to pure differential privacy. Prior algorithms for this…

Data Structures and Algorithms · Computer Science 2022-06-06 Samuel B. Hopkins , Gautam Kamath , Mahbod Majid

Let X Nv(0, {\Lambda}) be a normal vector in v dimensions, where {\Lambda} is diagonal. With reference to the truncated distribution of X on the interior of a v-dimensional Euclidean ball, we completely prove a variance inequality and a…

Statistics Theory · Mathematics 2013-11-26 Rahul Mukerjee , S. H. Ong

It was shown that when one disposes of a parametric information of the truncation distribution, the semiparametric estimator of the distribution function for truncated data (Wang, 1989) is more efficient than the nonparametric one. On the…

Statistics Theory · Mathematics 2021-06-03 Saida Mancer , Abdelhakim Necir , Souad Benchaira

Randomized algorithms depend on accurate sampling from probability distributions, as their correctness and performance hinge on the quality of the generated samples. However, even for common distributions like Binomial, exact sampling is…

Computation · Statistics 2025-06-17 Uddalok Sarkar , Sourav Chakraborty , Kuldeep S. Meel

We study the problem of high-dimensional linear regression in a robust model where an $\epsilon$-fraction of the samples can be adversarially corrupted. We focus on the fundamental setting where the covariates of the uncorrupted samples are…

Machine Learning · Computer Science 2018-06-04 Ilias Diakonikolas , Weihao Kong , Alistair Stewart

We give the first polynomial-time, polynomial-sample, differentially private estimator for the mean and covariance of an arbitrary Gaussian distribution $\mathcal{N}(\mu,\Sigma)$ in $\mathbb{R}^d$. All previous estimators are either…

Machine Learning · Statistics 2022-02-15 Gautam Kamath , Argyris Mouzakis , Vikrant Singhal , Thomas Steinke , Jonathan Ullman

We study the convergence rate of randomly truncated stochastic algorithms, which consist in the truncation of the standard Robbins-Monro procedure on an increasing sequence of compact sets. Such a truncation is often required in practice to…

Probability · Mathematics 2010-04-08 Jérôme Lelong

We study the convergence rate of randomly truncated stochastic algorithms, which consist in the truncation of the standard Robbins-Monro procedure on an increasing sequence of compact sets. Such a truncation is often required in practice to…

Probability · Mathematics 2010-03-23 Jérôme Lelong

We study the problem of high-dimensional sparse mean estimation in the presence of an $\epsilon$-fraction of adversarial outliers. Prior work obtained sample and computationally efficient algorithms for this task for identity-covariance…

Data Structures and Algorithms · Computer Science 2024-07-08 Ilias Diakonikolas , Daniel M. Kane , Sushrut Karmalkar , Ankit Pensia , Thanasis Pittas

In prevalent cohort studies where subjects are recruited at a cross-section, the time to an event may be subject to length-biased sampling, with the observed data being either the forward recurrence time, or the backward recurrence time, or…

Statistics Theory · Mathematics 2019-04-05 Pourab Roy , Jason P. Fine , Michael R. Kosorok

We propose a method to efficiently integrate truncated probability densities. The method uses Markov chain Monte Carlo method to sample from a probability density matching the function being integrated. The required normalisation or…

Computation · Statistics 2013-12-10 A. John Arul , Kannan Iyer

Elliptical slice sampling, when adapted to linearly truncated multivariate normal distributions, is a rejection-free Markov chain Monte Carlo method. At its core, it requires analytically constructing an ellipse-polytope intersection. The…

Machine Learning · Computer Science 2024-07-16 Kaiwen Wu , Jacob R. Gardner

We present a new version of the truncated harmonic mean estimator (THAMES) for univariate or multivariate mixture models. The estimator computes the marginal likelihood from Markov chain Monte Carlo (MCMC) samples, is consistent,…

We revisit the problem of estimating the mean of a real-valued distribution, presenting a novel estimator with sub-Gaussian convergence: intuitively, "our estimator, on any distribution, is as accurate as the sample mean is for the Gaussian…

Statistics Theory · Mathematics 2020-11-18 Jasper C. H. Lee , Paul Valiant

The analysis of a truncated sample can be hindered by censoring. Survival information may be lost to follow-up or the birthdate may be missing. The data can still be modeled as a truncated point process and it is close to a Poisson process,…

Methodology · Statistics 2025-08-12 Fiete Sieg , Anne-Marie Toparkus , Rafael Weissbach

The drift sequential parameter estimation problems for the Cox-Ingersoll-Ross (CIR) processes under the limited duration of observation are studied. Truncated sequential estimation methods for both scalar and {two}-dimensional parameter…

Statistics Theory · Mathematics 2025-04-08 Mohamed Ben Alaya , Thi-Bao Trâm Ngô , Serguei Pergamenchtchikov

Many statistical estimators are defined as the fixed point of a data-dependent operator, with estimators based on minimizing a cost function being an important special case. The limiting performance of such estimators depends on the…

Machine Learning · Computer Science 2022-03-22 Nhat Ho , Koulik Khamaru , Raaz Dwivedi , Martin J. Wainwright , Michael I. Jordan , Bin Yu

Statistical and machine-learning algorithms are frequently applied to high-dimensional data. In many of these applications data is scarce, and often much more costly than computation time. We provide the first sample-efficient…

Machine Learning · Computer Science 2014-02-20 Jayadev Acharya , Ashkan Jafarpour , Alon Orlitsky , Ananda Theertha Suresh

This paper develops recurrence relations for integrals that relate the density of multivariate extended skew-normal (ESN) distribution, including the well-known skew-normal (SN) distribution introduced by Azzalini and Dalla-Valle (1996) and…

Statistics Theory · Mathematics 2020-09-29 Christian E. Galarza , Larissa A. Matos , Dipak K. Dey , Victor H. Lachos
‹ Prev 1 3 4 5 6 7 10 Next ›