English
Related papers

Related papers: Cluster variation - Pade` approximants method for …

200 papers

Cluster analysis requires many decisions: the clustering method and the implied reference model, the number of clusters and, often, several hyper-parameters and algorithms' tunings. In practice, one produces several partitions, and a final…

Machine Learning · Statistics 2023-08-14 Luca Coraggio , Pietro Coretto

In recent years new application areas have emerged in which one aims to capture the geometry of objects by means of three-dimensional point clouds. Often the obtained data consist of a dense sampling of the object's surface, containing many…

Numerical Analysis · Mathematics 2019-10-01 Daniel Tenbrinck , Fjedor Gaede , Martin Burger

Crowdsourced, or human computation based clustering algorithms usually rely on relative distance comparisons, as these are easier to elicit from human workers than absolute distance information. A relative distance comparison is a statement…

Data Structures and Algorithms · Computer Science 2017-09-26 Antti Ukkonen

Determining the number of clusters is a fundamental issue in data clustering. Several algorithms have been proposed, including centroid-based algorithms using the Euclidean distance and model-based algorithms using a mixture of probability…

Machine Learning · Computer Science 2024-07-30 Ryosuke Motegi , Yoichi Seki

Machine learning has been successfully applied to identify phases and phase transitions in condensed matter systems. However, quantitative characterization of the critical fluctuations near phase transitions is lacking. In this study we…

Disordered Systems and Neural Networks · Physics 2019-03-19 Zhenyu Li , Mingxing Luo , Xin Wan

We propose a new effective cluster algorithm of tuning the critical point automatically, which is an extended version of Swendsen-Wang algorithm. We change the probability of connecting spins of the same type, $p = 1 - e^{- J/ k_BT}$, in…

Statistical Mechanics · Physics 2009-10-31 Yusuke Tomita , Yutaka Okabe

We introduce a new method for performing clustering with the aim of fitting clusters with different scatters and weights. It is designed by allowing to handle a proportion $\alpha$ of contaminating data to guarantee the robustness of the…

Statistics Theory · Mathematics 2008-12-18 Luis A. García-Escudero , Alfonso Gordaliza , Carlos Matrán , Agustin Mayo-Iscar

Gaussian variational approximations are widely used for summarizing posterior distributions in Bayesian models, especially in high-dimensional settings. However, a drawback of such approximations is the inability to capture skewness or more…

Methodology · Statistics 2026-04-02 Lucas Kock , Linda S. L. Tan , Prateek Bansal , David J. Nott

The domain of explainable AI is of interest in all Machine Learning fields, and it is all the more important in clustering, an unsupervised task whose result must be validated by a domain expert. We aim at finding a clustering that has high…

Artificial Intelligence · Computer Science 2024-03-28 Mathieu Guilbert , Christel Vrain , Thi-Bich-Hanh Dao

We propose a new approach for scaling prior to cluster analysis based on the concept of pooled variance. Unlike available scaling procedures such as the standard deviation and the range, our proposed scale avoids dampening the beneficial…

Methodology · Statistics 2020-07-28 Jakob Raymaekers , Ruben H. Zamar

It is often assumed that for treating numerical (or experimental) data on continuous transitions the formal analysis derived from the Renormalization Group Theory can only be applied over a narrow temperature range, the "critical region";…

Statistical Mechanics · Physics 2015-05-20 I. A. Campbell , P. H. Lundow

Growth mixture models are an important tool for detecting group structure in repeated measures data. Unlike traditional clustering methods, they explicitly model the repeat measurements on observations, and the statistical framework they…

Methodology · Statistics 2017-10-20 Abby Flynt , Nema Dean

Bayesian hierarchical clustering (BHC) is an agglomerative clustering method, where a probabilistic model is defined and its marginal likelihoods are evaluated to decide which clusters to merge. While BHC provides a few advantages over…

Machine Learning · Statistics 2015-06-04 Juho Lee , Seungjin Choi

We compute the structure factor of the $J_1$-$J_2$ Ising model in an external field on the square lattice within the Cluster Variation Method. We use a four point plaquette approximation, which is the minimal one able to capture phases with…

Statistical Mechanics · Physics 2016-10-19 Alejandra I. Guerrero , Daniel A. Stariolo

We present a clustering method and provide a theoretical analysis and an explanation to a phenomenon encountered in the applied statistical literature since the 1990's. This phenomenon is the natural adaptability of the order when using a…

Statistics Theory · Mathematics 2022-03-23 Thierry Dumont

The PAC-Bayesian approach is a powerful set of techniques to derive non- asymptotic risk bounds for random estimators. The corresponding optimal distribution of estimators, usually called the Gibbs posterior, is unfortunately intractable.…

Machine Learning · Statistics 2015-06-16 Pierre Alquier , James Ridgway , Nicolas Chopin

Classical peaks over threshold analysis is widely used for statistical modeling of sample extremes, and can be supplemented by a model for the sizes of clusters of exceedances. Under mild conditions a compound Poisson process model allows…

Applications · Statistics 2016-08-14 Mária Süveges , Anthony C. Davison

Clustering algorithms are one of the main analytical methods to detect patterns in unlabeled data. Existing clustering methods typically treat samples in a dataset as points in a metric space and compute distances to group together similar…

Machine Learning · Computer Science 2021-10-12 Tarek Naous , Srinjay Sarkar , Abubakar Abid , James Zou

We apply extensive Monte Carlo simulations to study the probability distribution $P(m)$ of the order parameter $m$ for the simple cubic Ising model with periodic boundary condition at the transition point. Sampling is performed with the…

Computational Physics · Physics 2020-03-18 Jiahao Xu , Alan M. Ferrenberg , David P. Landau

This paper deals with clustering methods based on adaptive distances for histogram data using a dynamic clustering algorithm. Histogram data describes individuals in terms of empirical distributions. These kind of data can be considered as…

Statistics Theory · Mathematics 2016-05-03 Antonio Irpino , Rosanna Verde , Francisco de AT De Carvalho