English
Related papers

Related papers: New Tests of Spatial Segregation Based on Nearest …

200 papers

Spatial clustering is a crucial field, finding universal use across criminology, pathology, and urban planning. However, most spatial clustering algorithms cannot pull information from nearby nodes and suffer performance drops when dealing…

Machine Learning · Computer Science 2025-03-12 Aidan Gao , Junhong Lin

Conditional randomization tests (CRTs) assess whether a variable $x$ is predictive of another variable $y$, having observed covariates $z$. CRTs require fitting a large number of predictive models, which is often computationally…

Methodology · Statistics 2023-04-12 Mukund Sudarshan , Aahlad Manas Puli , Wesley Tansey , Rajesh Ranganath

Sparse representation (SR) and collaborative representation (CR) have been successfully applied in many pattern classification tasks such as face recognition. In this paper, we propose a novel Non-negative Sparse and Collaborative…

Computer Vision and Pattern Recognition · Computer Science 2022-05-13 Jun Xu , Zhou Xu , Wangpeng An , Haoqian Wang , David Zhang

A key objective in spatial statistics is to simulate from the distribution of a spatial process at a selection of unobserved locations conditional on observations (i.e., a predictive distribution) to enable spatial prediction and…

Methodology · Statistics 2025-11-17 Julia Walchessen , Andrew Zammit-Mangion , Raphaël Huser , Mikael Kuusela

High dimensional data analysis for exploration and discovery includes three fundamental tasks: dimensionality reduction, clustering, and visualization. When the three associated tasks are done separately, as is often the case thus far,…

Machine Learning · Computer Science 2020-12-02 Stan Z. Li , Lirong Wu , Zelin Zang

Spatial clustering has been widely used for spatial data mining and knowledge discovery. An ideal multivariate spatial clustering should consider both spatial contiguity and aspatial attributes. Existing spatial clustering approaches may…

Machine Learning · Computer Science 2022-04-01 Yuhao Kang , Kunlin Wu , Song Gao , Ignavier Ng , Jinmeng Rao , Shan Ye , Fan Zhang , Teng Fei

Two-sample tests for multivariate data and especially for non-Euclidean data are not well explored. This paper presents a novel test statistic based on a similarity graph constructed on the pooled observations from the two samples. It can…

Methodology · Statistics 2024-08-12 Hao Chen , Jerome H. Friedman

Computing meaningful clusters of nodes is crucial to analyse large networks. In this paper, we apply new clustering methods to improve the computational time. We use the properties of the adjacency matrix to obtain better role extraction.…

Social and Information Networks · Computer Science 2017-02-22 Sibo Cheng , Adissa Laurent , Paul Van Dooren

We consider clustering based on significance tests for Gaussian Mixture Models (GMMs). Our starting point is the SigClust method developed by Liu et al. (2008), which introduces a test based on the k-means objective (with k = 2) to decide…

Methodology · Statistics 2019-10-08 Purvasha Chakravarti , Sivaraman Balakrishnan , Larry Wasserman

Testing independence among a number of (ultra) high-dimensional random samples is a fundamental and challenging problem. By arranging $n$ identically distributed $p$-dimensional random vectors into a $p \times n$ data matrix, we investigate…

Statistics Theory · Mathematics 2017-03-28 Xi Chen , Weidong Liu

A new method based on the rejection sampling for finding statistical tests is proposed. This method is conceptually intuitive, easy to implement, and applicable for arbitrary dimension. To illustrate its potential applicability, three…

Methodology · Statistics 2026-03-11 Markku Kuismin

Investigations of mass segregation are of vital interest for the understanding of the formation and dynamical evolution of stellar systems on a wide range of spatial scales. Our method is based on the minimum spanning tree (MST) that serves…

Astrophysics of Galaxies · Physics 2015-05-28 C. Olczak , R. Spurzem , Th. Henning

Conditional independence (CI) testing arises naturally in many scientific problems and applications domains. The goal of this problem is to investigate the conditional independence between a response variable $Y$ and another variable $X$,…

Methodology · Statistics 2025-10-07 Adel Javanmard , Mohammad Mehrabi

Cross-correlations between datasets are used in many different contexts in cosmological analyses. Recently, $k$-Nearest Neighbor Cumulative Distribution Functions ($k{\rm NN}$-${\rm CDF}$) were shown to be sensitive probes of cosmological…

Cosmology and Nongalactic Astrophysics · Physics 2021-04-28 Arka Banerjee , Tom Abel

Recent developed deep unsupervised methods allow us to jointly learn representation and cluster unlabelled data. These deep clustering methods mainly focus on the correlation among samples, e.g., selecting high precision pairs to gradually…

Computer Vision and Pattern Recognition · Computer Science 2019-08-13 Jianlong Wu , Keyu Long , Fei Wang , Chen Qian , Cheng Li , Zhouchen Lin , Hongbin Zha

We investigate clustering properties of dark matter halos and galaxies to search for optimal statistics and scales where possible departures from general relativity (GR) could be found. We use large N-body cosmological simulations to…

Cosmology and Nongalactic Astrophysics · Physics 2021-05-26 Jorge Enrique García-Farieta , Wojciech A. Hellwing , Suhani Gupta , Maciej Bilicki

Sample-level rankings are increasingly used in data-centric NLP for analysis, filtering, debugging, and curation, yet existing pipelines typically score training examples pointwise and rank them as if they were independent. This assumption…

Information Retrieval · Computer Science 2026-05-05 Xu Zheng , Feiyu Wu , Linhong Wu , Zhuocheng Wang , Hui Li

Rank correlations have found many innovative applications in the last decade. In particular, suitable rank correlations have been used for consistent tests of independence between pairs of random variables. Using ranks is especially…

Statistics Theory · Mathematics 2021-05-04 Hongjian Shi , Marc Hallin , Mathias Drton , Fang Han

Testing mutual independence for high-dimensional observations is a fundamental statistical challenge. Popular tests based on linear and simple rank correlations are known to be incapable of detecting non-linear, non-monotone relationships,…

Statistics Theory · Mathematics 2020-02-06 Mathias Drton , Fang Han , Hongjian Shi

The assumption of normality has underlain much of the development of statistics, including spatial statistics, and many tests have been proposed. In this work, we focus on the multivariate setting and first review the recent advances in…

Methodology · Statistics 2022-05-18 Wanfang Chen , Marc G. Genton
‹ Prev 1 4 5 6 7 8 10 Next ›