English
Related papers

Related papers: Limit theorems for unbounded cluster functionals o…

200 papers

We study the statistics of the maximum and minimum of a set of $N$ random variables whose dynamical and statistical properties fall within the scope of infinite ergodic theory. These non-stationary yet recurrent systems are described, in…

Statistical Mechanics · Physics 2026-03-09 Talia Baravi , Eli Barkai

This paper considers the problem of inference in cluster randomized experiments when cluster sizes are non-ignorable. Here, by a cluster randomized experiment, we mean one in which treatment is assigned at the cluster level. By…

Econometrics · Economics 2024-04-11 Federico Bugni , Ivan Canay , Azeem Shaikh , Max Tabord-Meehan

Understanding the complex structure of multivariate extremes is a major challenge in various fields from portfolio monitoring and environmental risk management to insurance. In the framework of multivariate Extreme Value Theory, a common…

Machine Learning · Statistics 2021-02-09 Hamid Jalalzai , Rémi Leluc

We consider the problem of clustering (or reconstruction) in the stochastic block model, in the regime where the average degree is constant. For the case of two clusters with equal sizes, recent results by Mossel, Neeman and Sly, and by…

Probability · Mathematics 2014-04-28 Joe Neeman , Praneeth Netrapalli

A motion of point vortices with periodic boundary conditions is studied by using Weierstrass zeta functions. Scattering and recoupling of a vortex pair by a third vortex becomes remarkable when the vortex density is large. Clustering of…

Fluid Dynamics · Physics 2007-05-23 Makoto Umeki

We consider the problem of clustering functional data while jointly selecting the most relevant features for classification. This problem has never been tackled before in the functional data context, and it requires a proper definition of…

Methodology · Statistics 2015-01-21 Davide Floriello , Valeria Vitelli

The core of the classical block maxima method consists of fitting an extreme value distribution to a sample of maxima over blocks extracted from an underlying series. In asymptotic theory, it is usually postulated that the block maxima are…

Statistics Theory · Mathematics 2014-05-09 Axel Bücher , Johan Segers

The aim of this paper is to develop tractable large deviation approximations for the empirical measure of a small noise diffusion. The starting point is the Freidlin-Wentzell theory, which shows how to approximate via a large deviation…

Probability · Mathematics 2021-01-11 Paul Dupuis , Guo-Jhen Wu

Consider a particle moving through a random medium, which consists of spherical obstacles, randomly distributed in R^d. The particle is accelerated by a constant external field; when colliding with an obstacle, the particle inelastically…

Probability · Mathematics 2007-05-23 Vladislav Vysotsky

We propose some axioms for hierarchical clustering of probability measures and investigate their ramifications. The basic idea is to let the user stipulate the clusters for some elementary measures. This is done without the need of any…

Machine Learning · Statistics 2016-05-24 Philipp Thomann , Ingo Steinwart , Nico Schmid

We are interested in a fragmentation process. We observe fragments frozen when their sizes are less than $\epsilon$ ($\epsilon$ > 0). Is is known ([BM05]) that the empirical measure of these fragments converges in law, under some…

Probability · Mathematics 2019-07-30 Sylvain Rubenthaler

Typically clustering algorithms provide clustering solutions with prespecified number of clusters. The lack of a priori knowledge on the true number of underlying clusters in the dataset makes it important to have a metric to compare the…

Machine Learning · Computer Science 2018-11-20 Amber Srivastava , Mayank Baranwal , Srinivasa Salapaka

Persistence is considered in diffusion--limited cluster--cluster aggregation, in one dimension and when the diffusion coefficient of a cluster depends on its size $s$ as $D(s) \sim s^\gamma$. The empty and filled site persistences are…

Statistical Mechanics · Physics 2016-08-16 E. K. O. Hellén , M. J. Alava

The Cluster-cluster model was introduced by Meakin et al in 1984. Each $x\in \mathbb{Z}^d$ starts with a cluster of size 1 with probability $p \in (0,1]$ independently. Each cluster $C$ performs a continuous-time SRW with rate…

Probability · Mathematics 2025-07-08 Noam Berger , Eviatar B. Procaccia , Daniel Sharon

Clustering methods must be tailored to the dataset it operates on, as there is no objective or universal definition of ``cluster,'' but nevertheless arbitrariness in the clustering method must be minimized. This paper develops a…

Information Theory · Computer Science 2024-05-03 Brian Weber

We consider block codes whose rate converges to the channel capacity with increasing block length at a certain speed and examine the best possible decay of the probability of error. We prove that a moderate deviation principle holds for all…

Information Theory · Computer Science 2015-03-20 Yucel Altug , Aaron B. Wagner

Dirichlet process mixtures are flexible non-parametric models, particularly suited to density estimation and probabilistic clustering. In this work we study the posterior distribution induced by Dirichlet process mixtures as the sample size…

Statistics Theory · Mathematics 2022-11-29 Filippo Ascolani , Antonio Lijoi , Giovanni Rebaudo , Giacomo Zanella

The one-dimensional contact process is analyzed by a cluster approximation. In this approach, the hierarchy of rate equations for the densities of finite length empty intervals are truncated under the assumption that adjacent intervals are…

Condensed Matter · Physics 2009-10-22 E. Ben-Naim , P. L. Krapivsky

We propose a model-based clustering algorithm for a general class of functional data for which the components could be curves or images. The random functional data realizations could be measured with error at discrete, and possibly random,…

Machine Learning · Statistics 2022-03-14 Steven Golovkine , Nicolas Klutchnikoff , Valentin Patilea

We present two methods for detecting patterns and clusters in high dimensional time-dependent functional data. Our methods are based on wavelet-based similarity measures, since wavelets are well suited for identifying highly discriminant…

Methodology · Statistics 2013-02-15 Anestis Antoniadis , Xavier Brossat , Jairo Cugliari , Jean-Michel Poggi