English
Related papers

Related papers: Regularity in the distribution of superclusters?

200 papers

We investigate the clustering properties and close neighbour counts for galaxies with different types of bulges and stellar masses. We select samples of "classical" and "pseudo" bulges, as well as "bulge-less" disk galaxies, based on the…

Astrophysics of Galaxies · Physics 2019-03-15 Lan Wang , Lixin Wang , Cheng Li , Jian Hu , Houjun Mo , Huiyuan Wang

Spectral clustering is one of the most prominent clustering approaches. The distance-based similarity is the most widely used method for spectral clustering. However, people have already noticed that this is not suitable for multi-scale…

Machine Learning · Computer Science 2020-09-11 Hengrui Wang , Yubo Zhang , Mingzhi Chen , Tong Yang

Jeong and Steinhardt (JS) implement local rules by selecting the sub-ensemble of tilings which have the maximum occurrences of a chosen pattern (``cluster'') C It is unknown how to prove that a given C implies a given sub-ensemble;…

Condensed Matter · Physics 2007-05-23 C. L. Henley

We propose some axioms for hierarchical clustering of probability measures and investigate their ramifications. The basic idea is to let the user stipulate the clusters for some elementary measures. This is done without the need of any…

Machine Learning · Statistics 2016-05-24 Philipp Thomann , Ingo Steinwart , Nico Schmid

Using the APM cluster distribution we find interesting alignment effects: (1) Cluster substructure is strongly correlated with the tendency of clusters to be aligned with their nearest neighbour and in general with the nearby clusters that…

Astrophysics · Physics 2017-01-25 Manolis Plionis

A popular method for selecting the number of clusters is based on stability arguments: one chooses the number of clusters such that the corresponding clustering results are "most stable". In recent years, a series of papers has analyzed the…

Machine Learning · Statistics 2010-07-08 Ulrike von Luxburg

As a kind of basic machine learning method, clustering algorithms group data points into different categories based on their similarity or distribution. We present a clustering algorithm by finding hyper-planes to distinguish the data…

Computer Vision and Pattern Recognition · Computer Science 2020-04-28 Luhong Diao , Jinying Gao1 , Manman Deng

The realization that most stars form in clusters, raises the question of whether star/planet formation are influenced by the cluster environment. The stellar density in the most prevalent clusters is the key factor here. Whether dominant…

Astrophysics of Galaxies · Physics 2015-06-11 S. Pfalzner , T. Kaczmarek , C. Olczak

Numerous papers ask how difficult it is to cluster data. We suggest that the more relevant and interesting question is how difficult it is to cluster data sets {\em that can be clustered well}. More generally, despite the ubiquity and the…

Machine Learning · Computer Science 2012-05-23 Amit Daniely , Nati Linial , Michael Saks

A measure of distance between two clusterings has important applications, including clustering validation and ensemble clustering. Generally, such distance measure provides navigation through the space of possible clusterings. Mostly used…

Social and Information Networks · Computer Science 2015-09-01 Reihaneh Rabbany , Osmar R. Zaïane

We introduce a clustering coefficient for nondirected and directed hypergraphs, which we call the quad clustering coefficient. We determine the average quad clustering coefficient and its distribution in real-world hypergraphs and compare…

Physics and Society · Physics 2024-04-08 Gyeong-Gyun Ha , Izaak Neri , Alessia Annibale

Consistency is a key property of all statistical procedures analyzing randomly sampled data. Surprisingly, despite decades of work, little is known about consistency of most clustering algorithms. In this paper we investigate consistency of…

Statistics Theory · Mathematics 2008-12-18 Ulrike von Luxburg , Mikhail Belkin , Olivier Bousquet

The clustering coefficient quantifies how well connected are the neighbors of a vertex in a graph. In real networks it decreases with the vertex degree, which has been taken as a signature of the network hierarchical structure. Here we show…

Statistical Mechanics · Physics 2007-05-23 Sara Nadiv Soffer , Alexei Vazquez

We describe the structure of the graphs with the smallest average distance and the largest average clustering given their order and size. There is usually a unique graph with the largest average clustering, which at the same time has the…

Molecular Networks · Quantitative Biology 2010-07-28 Dionysios Barmpoutis , Richard M. Murray

We develop a general theory for estimating the probability that a galaxy cluster of a given shape exists. The theory is based on the observed result that the distribution of galaxies is very close to quasi-equilibrium, in both its linear…

Cosmology and Nongalactic Astrophysics · Physics 2012-01-11 Abel Yang , William C. Saslaw

Datasets in high-dimension do not typically form clusters in their original space; the issue is worse when the number of points in the dataset is small. We propose a low-computation method to find statistically significant clustering…

Machine Learning · Statistics 2020-08-24 Alden Bradford , Tarun Yellamraju , Mireille Boutin

Genome wide comparisons between enteric bacteria yield large sets of conserved putative regulatory sites on a gene by gene basis that need to be clustered into regulons. Using the assumption that regulatory sites can be represented as…

Biological Physics · Physics 2009-11-07 Erik van Nimwegen , Mihaela Zavolan , Nikolaus Rajewsky , Eric D. Siggia

Let $X,X_1,X_2,\ldots$ be i.i.d. mean zero random vectors with values in a separable Banach space $B$, $S_n=X_1+\cdots+X_n$ for $n\ge1$, and assume $\{c_n:n\ge1\}$ is a suitably regular sequence of constants. Furthermore, let…

Probability · Mathematics 2014-03-28 Uwe Einmahl , Jim Kuelbs

Mixture model-based frameworks are very popular for statistical inference in clustering. While convenient for producing probabilistic estimates of cluster assignments and uncertainty, they are prone to misspecification, which can lead to…

Statistics Theory · Mathematics 2026-05-15 Yu Zheng , Leo L. Duan , Arkaprava Roy

A random network model which allows for tunable, quite general forms of clustering, degree correlation and degree distribution is defined. The model is an extension of the configuration model, in which stubs (half-edges) are paired to form…

Probability · Mathematics 2012-07-31 Frank Ball , Tom Britton , David Sirl