English
Related papers

Related papers: Fast rates for noisy clustering

200 papers

The complexity of the Quicksort algorithm is usually measured by the number of key comparisons used during its execution. When operating on a list of $n$ data, permuted uniformly at random, the appropriately normalized complexity $Y_n$ is…

Probability · Mathematics 2013-01-25 Ralph Neininger

Optimal estimation of a coin's bias using noisy data is surprisingly different from the same problem with noiseless data. We study this problem using entropy risk to quantify estimators' accuracy. We generalize the "add Beta" estimators…

Statistics Theory · Mathematics 2015-03-19 Christopher Ferrie , Robin Blume-Kohout

In this paper, we study the statistical properties of kernel $k$-means and obtain a nearly optimal excess clustering risk bound, substantially improving the state-of-art bounds in the existing clustering risk analyses. We further analyze…

Machine Learning · Computer Science 2020-05-15 Yong Liu , Lizhong Ding , Weiping Wang

In this paper, we review state-of-the-art methods for feature selection in statistics with an application-oriented eye. Indeed, sparsity is a valuable property and the profusion of research on the topic might have provided little guidance…

Methodology · Statistics 2021-11-08 Dimitris Bertsimas , Jean Pauphilet , Bart Van Parys

This paper addresses signal denoising when large-amplitude coefficients form clusters (groups). The L1-norm and other separable sparsity models do not capture the tendency of coefficients to cluster (group sparsity). This work develops an…

Computer Vision and Pattern Recognition · Computer Science 2017-02-21 Po-Yu Chen , Ivan W. Selesnick

We study the minimization of fixed-degree polynomials over the simplex. This problem is well-known to be NP-hard, as it contains the maximum stable set problem in graph theory as a special case. In this paper, we consider a rational…

Optimization and Control · Mathematics 2014-07-09 Etienne de Klerk , Monique Laurent , Zhao Sun

We give overcrowding estimates for the Sine_beta process, the bulk point process limit of the Gaussian beta-ensemble. We show that the probability of having at least n points in a fixed interval is given by $e^{-\frac{\beta}{2} n^2…

Probability · Mathematics 2015-06-24 Diane Holcomb , Benedek Valkó

We develop a methodology for conducting inference on extreme quantiles of unobserved individual heterogeneity (e.g., heterogeneous coefficients, treatment effects) in panel data and meta-analysis settings. Inference is challenging in such…

Econometrics · Economics 2026-02-04 Vladislav Morozov

We address the problem of classification when data are collected from two samples with measurement errors. This problem turns to be an inverse problem and requires a specific treatment. In this context, we investigate the minimax rates of…

Statistics Theory · Mathematics 2013-07-15 Sébastien Loustau , Clément Marteau

In high-dimensional statistical inference, sparsity regularizations have shown advantages in consistency and convergence rates for coefficient estimation. We consider a generalized version of Sparse-Group Lasso which captures both…

Machine Learning · Statistics 2020-08-12 Xinyu Zhang

In this paper we propose a modified version of the simulated annealing algorithm for solving a stochastic global optimization problem. More precisely, we address the problem of finding a global minimizer of a function with noisy…

Machine Learning · Statistics 2017-03-02 Clément Bouttier , Ioana Gavra

The group testing problem is concerned with identifying a small set of infected individuals in a large population. At our disposal is a testing procedure that allows us to test several individuals together. In an idealized setting, a test…

Information Theory · Computer Science 2023-09-19 Oliver Gebhard , Oliver Johnson , Philipp Loick , Maurice Rolvien

The seminal paper by Mazumdar and Saha \cite{MS17a} introduced an extensive line of work on clustering with noisy queries. Yet, despite significant progress on the problem, the proposed methods depend crucially on knowing the exact…

Machine Learning · Computer Science 2022-07-22 Alberto Del Pia , Mingchen Ma , Christos Tzamos

Quantum metrology protocols allow to surpass precision limits typical to classical statistics. However, in recent years, no-go theorems have been formulated, which state that typical forms of uncorrelated noise can constrain the quantum…

Quantum Physics · Physics 2016-03-30 Andrea Smirne , Jan Kolodynski , Susana F. Huelga , Rafal Demkowicz-Dobrzanski

Let $X$ and $Y$ be two independent identically distributed random variables with density $p(x)$ and $Z=\alpha X+\beta Y$ for some constants $\alpha>0$ and $\beta>0$. We consider the problem of estimating $p(x)$ by means of the samples from…

Statistics Theory · Mathematics 2007-06-13 Denis Belomestny

We consider the problem of Bayesian optimization of a one-dimensional Brownian motion in which the $T$ adaptively chosen observations are corrupted by Gaussian noise. We show that as the smallest possible expected cumulative regret and the…

Machine Learning · Computer Science 2022-01-19 Zexin Wang , Vincent Y. F. Tan , Jonathan Scarlett

We analyze the noise in macro-particle methods used in plasma physics and fluid dynamics, leading to approaches for minimizing the total error, focusing on electrostatic models in one dimension. We describe kernel density estimation for…

Plasma Physics · Physics 2021-06-30 E. G. Evstatiev , J. M. Finn , B. A. Shadwick , N. Hengartner

The subject of this paper is the estimation of a probability measure on ${\mathbb R}^d$ from data observed with an additive noise, under the Wasserstein metric of order $p$ (with $p\geq 1$). We assume that the distribution of the errors is…

Statistics Theory · Mathematics 2013-07-22 Jérôme Dedecker , Bertrand Michel

We consider in this paper the problem of noisy 1-bit matrix completion under a general non-uniform sampling distribution using the max-norm as a convex relaxation for the rank. A max-norm constrained maximum likelihood estimate is…

Machine Learning · Statistics 2013-09-25 T. Tony Cai , Wen-Xin Zhou

Given a set of points, clustering consists of finding a partition of a point set into $k$ clusters such that the center to which a point is assigned is as close as possible. Most commonly, centers are points themselves, which leads to the…

Machine Learning · Computer Science 2023-10-16 Maria Sofia Bucarelli , Matilde Fjeldsø Larsen , Chris Schwiegelshohn , Mads Bech Toftrup
‹ Prev 1 4 5 6 7 8 10 Next ›