中文
相关论文

相关论文: Comparative Bi-stochastizations and Associated Clu…

200 篇论文

To control for multiscale effects in networks, one can transform the matrix of (in general) weighted, directed internodal flows to bistochastic (doubly-stochastic) form, using the iterative proportional fitting (Sinkhorn-Knopp) procedure,…

物理与社会 · 物理学 2015-03-13 Paul B. Slater

We extend to the beta-divergence (Itakura-Saito) case beta =0, the comparative bi-stochaticization analyses-previously conducted (arXiv:1208.3428) for the (Kullback-Leibler) beta=1 and (squared-Euclidean) beta = 2 cases -of the 3,107 -…

物理与社会 · 物理学 2012-10-26 Paul B. Slater

We present a number of variously rearranged matrix plots of the $3, 107 \times 3, 107$ 1995-2000 (asymmetric) intercounty migration table for the United States, principally in its bistochasticized form (all 3,107 row and column sums…

物理与社会 · 物理学 2012-07-03 Paul B. Slater

We have obtained a "hierarchical regionalization" of 3,107 county-level units of the United States based upon census-recorded 1995-2000 intercounty migration flows. The methodology employed was the two-stage (double-standardization and…

物理与社会 · 物理学 2019-04-09 Paul B. Slater

Most nations of the world periodically publish N x N origin-destination tables, recording the number of people who lived in geographic subdivision i at time t and j at t+1. We have developed and widely applied to such national tables and…

物理与社会 · 物理学 2012-07-03 Paul B. Slater

Data matrix having different sets of entities in its rows and columns are known as two mode data or affiliation data. Many practical problems require to find relationships between the two modes by simultaneously clustering the rows and…

数据结构与算法 · 计算机科学 2018-07-23 Briti Deb , Indrajit Mukherjee

Exponential families and mixture families are parametric probability models that can be geometrically studied as smooth statistical manifolds with respect to any statistical divergence like the Kullback-Leibler (KL) divergence or the…

机器学习 · 计算机科学 2018-03-21 Frank Nielsen , Gaëtan Hadjeres

The increasing availability of multiple network data has highlighted the need for statistical models for heterogeneous populations of networks. A convenient framework makes use of metrics to measure similarity between networks. In this…

统计方法学 · 统计学 2026-03-09 Francesco Barile , Simón Lunagómez , Bernardo Nipoti

Common clustering algorithms require multiple scans of all the data to achieve convergence, and this is prohibitive when large databases, with data arriving in streams, must be processed. Some algorithms to extend the popular K-means method…

应用统计 · 统计学 2017-12-22 Giacomo Aletti , Alessandra Micheletti

An efficient MCMC algorithm is presented to cluster the nodes of a network such that nodes with similar role in the network are clustered together. This is known as block-modelling or block-clustering. The model is the stochastic blockmodel…

统计计算 · 统计学 2012-11-09 Aaron F. McDaid , Thomas Brendan Murphy , Nial Friel , Neil J Hurley

Upon a matrix representation of a binary bipartite network, via the permutation invariance, a coupling geometry is computed to approximate the minimum energy macrostate of a network's system. Such a macrostate is supposed to constitute the…

应用统计 · 统计学 2018-02-02 Jiahui Guan , Hsieh Fushing

We present a method to discover differences between populations with respect to the spatial coherence of their oriented white matter microstructure in arbitrarily shaped white matter regions. This method is applied to diffusion MRI scans of…

This paper addresses the limitations of conventional vector quantization algorithms, particularly K-Means and its variant K-Means++, and investigates the Stochastic Quantization (SQ) algorithm as a scalable alternative for high-dimensional…

机器学习 · 计算机科学 2025-03-11 Anton Kozyriev , Vladimir Norkin

Clustering analysis by nonnegative low-rank approximations has achieved remarkable progress in the past decade. However, most approximation approaches in this direction are still restricted to matrix factorization. We propose a new low-rank…

机器学习 · 计算机科学 2012-06-22 Zhirong Yang , Erkki Oja

Stochastic Block Models (SBMs) are a fundamental tool for community detection in network analysis. But little theoretical work exists on the statistical performance of Bayesian SBMs, especially when the community count is unknown. This…

统计理论 · 数学 2021-01-19 Sheng Jiang , Surya Tokdar

Biclustering, also known as co-clustering or two-way clustering, simultaneously partitions the rows and columns of a data matrix to reveal submatrices with coherent patterns. Incorporating background knowledge into clustering to enhance…

最优化与控制 · 数学 2026-02-24 Antonio M. Sudoso

We consider the problem of clustering functional data according to their covariance structure. We contribute a soft clustering methodology based on the Wasserstein-Procrustes distance, where the in-between cluster variability is penalised…

统计方法学 · 统计学 2022-12-29 V. Masarotto , G. Masarotto

Co-clustering simultaneously clusters rows and columns, revealing more fine-grained groups. However, existing co-clustering methods suffer from poor scalability and cannot handle large-scale data. This paper presents a novel and scalable…

分布式、并行与集群计算 · 计算机科学 2025-03-20 Zihan Wu , Zhaoke Huang , Hong Yan

Cut-based directed graph (digraph) clustering often focuses on finding dense within-cluster or sparse between-cluster connections, similar to cut-based undirected graph clustering methods. In contrast, for flow-based clusterings the edges…

机器学习 · 计算机科学 2022-03-04 Koby Hayashi , Sinan G. Aksoy , Haesun Park

With the development of Big data technology, data analysis has become increasingly important. Traditional clustering algorithms such as K-means are highly sensitive to the initial centroid selection and perform poorly on non-convex…

机器学习 · 计算机科学 2023-07-28 Ying Xiao , Hou-biao Li , Yu-pu Zhang
‹ 上一页 1 2 3 10 下一页 ›