English
Related papers

Related papers: Two-sample test of sparse stochastic block models

200 papers

Statistical techniques are used in all branches of science to determine the feasibility of quantitative hypotheses. One of the most basic applications of statistical techniques in comparative analysis is the test of equality of two…

Methodology · Statistics 2018-05-01 Ayanendranath Basu , Abhijit Mandal , Nirian Martin , Leandro Pardo

Assume that we have a random sample from an absolutely continuous distribution (univariate, or multivariate) with a known functional form and some unknown parameters. In this paper, we have studied several parametric tests based on…

Statistics Theory · Mathematics 2024-05-14 Rahul Singh , Neeraj Misra

The stochastic block model (SBM) is a widely used framework for community detection in networks, where the network structure is typically represented by an adjacency matrix. However, conventional SBMs are not directly applicable to an…

Machine Learning · Statistics 2023-10-18 Jie Jian , Mu Zhu , Peijun Sang

Parametric hypothesis testing associated with two independent samples arises frequently in several applications in biology, medical sciences, epidemiology, reliability and many more. In this paper, we propose robust Wald-type tests for…

Methodology · Statistics 2019-05-09 Abhik Ghosh , Nirian Martin , Ayanendranath Basu , Leandro Pardo

Over the past decades, various methods for comparing the means of two log-normal have been proposed. Some of them are differing in terms of how the statistic test adjust to accept or to reject the null hypothesis. In this study, a new…

Statistics Theory · Mathematics 2014-05-20 Kamel Abdollahnezhad , M. Babanezhad , Ali Akbar Jafari

Two-sample hypothesis testing for random graphs arises naturally in neuroscience, social networks, and machine learning. In this paper, we consider a semiparametric problem of two-sample hypothesis testing for a class of latent position…

Methodology · Statistics 2015-06-19 Minh Tang , Avanti Athreya , Daniel L. Sussman , Vince Lyzinski , Carey E. Priebe

Testing the homogeneity between two samples of functional data is an important task. While this is feasible for intensely measured functional data, we explain why it is challenging for sparsely measured functional data and show what can be…

Methodology · Statistics 2022-07-05 Changbo Zhu , Jane-Ling Wang

In this paper we extend our previous work on the stochastic block model, a commonly used generative model for social and biological networks, and the problem of inferring functional groups or communities from the topology of the network. We…

Statistical Mechanics · Physics 2013-05-09 Aurelien Decelle , Florent Krzakala , Cristopher Moore , Lenka Zdeborová

A fundamental problem in network data analysis is to test Erd\"{o}s-R\'{e}nyi model $\mathcal{G}\left(n,\frac{a+b}{2n}\right)$ versus a bisection stochastic block model $\mathcal{G}\left(n,\frac{a}{n},\frac{b}{n}\right)$, where $a,b>0$ are…

Methodology · Statistics 2018-11-26 Mingao Yuan , Yang Feng , Zuofeng Shang

Community detection is an important task in network analysis, in which we aim to learn a network partition that groups together vertices with similar community-level connectivity patterns. By finding such groups of vertices with similar…

Machine Learning · Statistics 2015-05-25 Christopher Aicher , Abigail Z. Jacobs , Aaron Clauset

The problem of community detection receives great attention in recent years. Many methods have been proposed to discover communities in networks. In this paper, we propose a Gaussian stochastic blockmodel that uses Gaussian distributions to…

Social and Information Networks · Computer Science 2015-03-19 Junhao Zhang , Tongfei Chen , Junfeng Hu

We consider the problem of estimating the number of communities in a weighted balanced Stochastic Block Model. We construct hypothesis tests based on semidefinite programming and with a statistic coming from a GOE matrix to distinguish…

Statistics Theory · Mathematics 2025-02-25 Deborah Oliveira , Andressa Cerqueira , Roberto Oliveira

Network data is a major object data type that has been widely collected or derived from common sources such as brain imaging. Such data contains numeric, topological, and geometrical information, and may be necessarily considered in certain…

Methodology · Statistics 2021-06-29 Han Feng , Xing Qiu , Hongyu Miao

We present new families of goodness-of-fit tests of uniformity on a full-dimensional set $W\subset\R^d$ based on statistics related to edge lengths of random geometric graphs. Asymptotic normality of these statistics is proven under the…

Statistics Theory · Mathematics 2020-07-20 Bruno Ebner , Franz Nestmann , Matthias Schulte

Community identification in a network is an important problem in fields such as social science, neuroscience, and genetics. Over the past decade, stochastic block models (SBMs) have emerged as a popular statistical framework for this…

Statistics Theory · Mathematics 2018-10-02 Min Xu , Varun Jog , Po-Ling Loh

So-called linear rank statistics provide a means for distribution-free (even in finite samples), yet highly flexible, two-sample testing in the setting of univariate random variables. Their flexibility derives from a choice of weights that…

Methodology · Statistics 2023-10-03 Dan D. Erdmann-Pham

Random graph models with community structure have been studied extensively in the literature. For both the problems of detecting and recovering community structure, an interesting landscape of statistical and computational phase transitions…

Statistics Theory · Mathematics 2025-06-30 Cynthia Rush , Fiona Skerman , Alexander S. Wein , Dana Yang

We propose a valid and consistent test for the hypothesis that two latent distance random graphs on the same vertex set have the same generating latent positions, up to some unidentifiable similarity transformations. Our test statistic is…

Methodology · Statistics 2020-08-04 Yiran Wang , Minh Tang , Soumendra Nath Lahiri

A non parametric method based on the empirical likelihood is proposed for detecting the change in the coefficients of high-dimensional linear model where the number of model variables may increase as the sample size increases. This amounts…

Statistics Theory · Mathematics 2015-06-22 Gabriela Ciuperca , Zahraa Salloum

A variety of machine learning tasks---e.g., matrix factorization, topic modelling, and feature allocation---can be viewed as learning the parameters of a probability distribution over bipartite graphs. Recently, a new class of models for…

Machine Learning · Statistics 2017-12-07 Victor Veitch , Ekansh Sharma , Zacharie Naulet , Daniel M. Roy
‹ Prev 1 3 4 5 6 7 10 Next ›