English
Related papers

Related papers: Hypothesis Testing for Two Sample Comparison of Ne…

200 papers

Data objects taking value in a general metric space have become increasingly common in modern data analysis. In this paper, we study two important statistical inference problems, namely, two-sample testing and change-point detection, for…

Methodology · Statistics 2023-07-11 Feiyu Jiang , Changbo Zhu , Xiaofeng Shao

In this paper, we address the problem of two-sample testing in the presence of missing data under a variety of missingness mechanisms. Our focus is on the well-known energy distance-based two-sample test. In addition to the standard…

Methodology · Statistics 2025-08-18 Danijel G. Aleksić , Bojana Milošević

Nonparametric two sample or homogeneity testing is a decision theoretic problem that involves identifying differences between two random variables without making parametric assumptions about their underlying distributions. The literature is…

Statistics Theory · Mathematics 2015-10-14 Aaditya Ramdas , Nicolas Garcia , Marco Cuturi

Data-driven brain parcellations aim to provide a more accurate representation of an individual's functional connectivity, since they are able to capture individual variability that arises due to development or disease. This renders…

Neurons and Cognition · Quantitative Biology 2017-03-30 Sofia Ira Ktena , Salim Arslan , Sarah Parisot , Daniel Rueckert

Hypothesis testing in high dimensional data is a notoriously difficult problem without direct access to competing models' likelihood functions. This paper argues that statistical divergences can be used to quantify the difference between…

Data Analysis, Statistics and Probability · Physics 2024-08-02 Jeremy J. H. Wilkinson , Christopher G. Lester

Deep Learning performs well when training data densely covers the experience space. For complex problems this makes data collection prohibitively expensive. We propose to intelligently select samples when constructing data sets in order to…

Computer Vision and Pattern Recognition · Computer Science 2020-04-01 Mark Philip Philipsen , Thomas Baltzer Moeslund

The surge in digitized text data requires reliable inferential methods on observed textual patterns. This article proposes a novel two-sample text test for comparing similarity between two groups of documents. The hypothesis is whether the…

Machine Learning · Statistics 2025-05-09 Jingbin Xu , Chen Qian , Meimei Liu , Feng Guo

We investigate a network model based on an infinite regular square lattice embedded in the Euclidean plane where the node connection probability is given by the geometrical distance of nodes. We show that the degree distribution in the…

Physics and Society · Physics 2008-06-23 Matus Medo , Jan Smrek

Most social, technological and biological networks are embedded in a finite dimensional space, and the distance between two nodes influences the likelihood that they link to each other. Indeed, in social systems, the chance that two…

Physics and Society · Physics 2018-06-27 Paul Balister , Chaoming Song , Oliver Riordan , Bela Bollobas , Albert-Laszlo Barabasi

The sampling method has been paid much attention in the field of complex network in general and statistical physics in particular. This paper presents two new sampling methods based on the perspective that a small part of vertices with high…

Physics and Society · Physics 2014-05-23 Luo Peng , Li Yongli , Wu Chong

Contacts between individuals play an important role in determining how infectious diseases spread. Various methods to gather data on such contacts co-exist, from surveys to wearable sensors. Comparisons of data obtained by different methods…

Physics and Society · Physics 2016-04-18 Julie Fournet , Alain Barrat

Multidimensional network data can have different levels of complexity, as nodes may be characterized by heterogeneous individual-specific features, which may vary across the networks. This paper introduces a class of models for…

Methodology · Statistics 2021-04-01 Silvia D'Angelo , Marco Alfò , Thomas Brendan Murphy

We consider the problem of testing positively dependent multiple hypotheses assuming that a prior information about the dependence structure is available. We propose two-step multiple comparisons procedures that exploit the prior…

Understanding the space of probability measures on a metric space equipped with a Wasserstein distance is one of the fundamental questions in mathematical analysis. The Wasserstein metric has received a lot of attention in the machine…

Machine Learning · Computer Science 2021-03-02 Arijit Sehanobish , Neal Ravindra , David van Dijk

Network models are widely used to represent relational information among interacting units and the structural implications of these relations. Recently, social network studies have focused a great deal of attention on random graph models of…

Applications · Statistics 2010-10-06 Mark S. Handcock , Krista J. Gile

Network node similarity measure has been paid particular attention in the field of statistical physics. In this paper, we utilize the concept of information and information loss to measure the node similarity. The whole model is based on…

Physics and Society · Physics 2014-03-19 Yongli Li , Peng Luo , Chong Wu

Identifying networks with similar characteristics in a given ensemble, or detecting pattern discontinuities in a temporal sequence of networks, are two examples of tasks that require an effective metric capable of quantifying network…

Social and Information Networks · Computer Science 2023-09-07 Carlo Piccardi

We study the problem of comparing a pair of geometric networks that may not be similarly defined, i.e., when they do not have one-to-one correspondences between their nodes and edges. Our motivating application is to compare power…

Computational Geometry · Computer Science 2025-06-10 Kostiantyn Lyman , Rounak Meyur , Bala Krishnamoorthy , Mahantesh Halappanavar

We consider the testing and estimation of change-points, locations where the distribution abruptly changes, in a sequence of multivariate or non-Euclidean observations. We study a nonparametric framework that utilizes similarity information…

Methodology · Statistics 2018-02-23 Lynna Chu , Hao Chen

In this work, we consider hypothesis testing and anomaly detection on datasets where each observation is a weighted network. Examples of such data include brain connectivity networks from fMRI flow data, or word co-occurrence counts for…

Machine Learning · Statistics 2018-09-10 Guilherme Gomes , Vinayak Rao , Jennifer Neville