English
Related papers

Related papers: Reducing Seed Bias in Respondent-Driven Sampling b…

200 papers

Mixture models are probabilistic models aimed at uncovering and representing latent subgroups within a population. In the realm of network data analysis, the latent subgroups of nodes are typically identified by their connectivity…

Methodology · Statistics 2020-05-27 Giacomo De Nicola , Benjamin Sischka , Göran Kauermann

Sparse-representation-based classification (SRC) has been widely studied and developed for various practical signal classification applications. However, the performance of a SRC-based method is degraded when both the training and test data…

Computer Vision and Pattern Recognition · Computer Science 2019-11-26 He-Feng Yin , Xiao-Jun Wu , Josef Kittler , Zhen-Hua Feng

With ever-increasing available data, predicting individuals' preferences and helping them locate the most relevant information has become a pressing need. Understanding and predicting preferences is also important from a fundamental point…

Physics and Society · Physics 2012-10-05 Roger Guimera , Alejandro Llorente , Esteban Moro , Marta Sales-Pardo

Worker recruitment is a crucial research problem in Mobile Crowd Sensing (MCS). While previous studies rely on a specified platform with a pre-assumed large user pool, this paper leverages the influenced propagation on the social network to…

Social and Information Networks · Computer Science 2018-05-23 Jiangtao Wang , Feng Wang , Yasha Wang , Daqing Zhang , Leye Wang , Zhaopeng Qiu

Network surveys of key populations at risk for HIV are an essential part of the effort to understand how the epidemic spreads and how it can be prevented. Estimation of population values from the sample data has been probematical, however,…

Applications · Statistics 2019-09-12 Steve Thompson

Analysis of social networks with limited data access is challenging for third parties. To address this challenge, a number of studies have developed algorithms that estimate properties of social networks via a simple random walk. However,…

Social and Information Networks · Computer Science 2023-05-23 Kazuki Nakajima , Kazuyuki Shudo

The proliferation of models for networks raises challenging problems of model selection: the data are sparse and globally dependent, and models are typically high-dimensional and have large numbers of latent variables. Together, these…

Social and Information Networks · Computer Science 2014-06-25 Xiaoran Yan , Cosma Rohilla Shalizi , Jacob E. Jensen , Florent Krzakala , Cristopher Moore , Lenka Zdeborova , Pan Zhang , Yaojia Zhu

Methods for ranking the importance of nodes in a network have a rich history in machine learning and across domains that analyze structured data. Recent work has evaluated these methods though the seed set expansion problem: given a subset…

Social and Information Networks · Computer Science 2017-05-04 Isabel Kloumann , Johan Ugander , Jon Kleinberg

In (\cite{zhang2014nonlinear,zhang2014nonlinear2}), we have viewed machine learning as a coding and dimensionality reduction problem, and further proposed a simple unsupervised dimensionality reduction method, entitled deep distributed…

Machine Learning · Computer Science 2015-01-29 Xiao-Lei Zhang

Online planning in Markov Decision Processes (MDPs) enables agents to make sequential decisions by simulating future trajectories from the current state, making it well-suited for large-scale or dynamic environments. Sample-based methods…

Artificial Intelligence · Computer Science 2025-09-22 Tamir Shazman , Idan Lev-Yehudi , Ron Benchetit , Vadim Indelman

This paper proposes a stochastic block model with dynamics where the population grows using preferential attachment. Nodes with higher weighted degree are more likely to recruit new nodes, and nodes always recruit nodes from their own…

Social and Information Networks · Computer Science 2024-01-15 Simina Brânzei , Nithish Kumar , Gireeja Ranade

We consider a machine learning setup where one training dataset is used to train multiple models across slightly different data distributions. This occurs when customized models are needed for various deployment environments. To reduce…

This paper proposes a distributed pseudo-likelihood method (DPL) to conveniently identify the community structure of large-scale networks. Specifically, we first propose a block-wise splitting method to divide large-scale network data into…

Methodology · Statistics 2024-11-05 Jiayi Deng , Danyang Huang , Bo Zhang

In pattern recognition, handling uncertainty is a critical challenge that significantly affects decision-making and classification accuracy. Dempster-Shafer Theory (DST) is an effective reasoning framework for addressing uncertainty, and…

Artificial Intelligence · Computer Science 2024-10-31 Juntao Xu , Tianxiang Zhan , Yong Deng

The boom of DL technology leads to massive DL models built and shared, which facilitates the acquisition and reuse of DL models. For a given task, we encounter multiple DL models available with the same functionality, which are considered…

Software Engineering · Computer Science 2021-03-10 Linghan Meng , Yanhui Li , Lin Chen , Zhi Wang , Di Wu , Yuming Zhou , Baowen Xu

Modeling relations between individuals is a classical question in social sciences and clustering individuals according to the observed patterns of interactions allows to uncover a latent structure in the data. Stochastic block model (SBM)…

Methodology · Statistics 2015-01-27 Pierre Barbillon , Sophie Donnet , Emmanuel Lazega , Avner Bar-Hen

Network-based clustering methods frequently require the number of communities to be specified \emph{a priori}. Moreover, most of the existing methods for estimating the number of communities assume the number of communities to be fixed and…

Methodology · Statistics 2022-01-14 Chetkar Jha , Mingyao Li , Ian Barnett

Distribution Regression (DR) on stochastic processes describes the learning task of regression on collections of time series. Path signatures, a technique prevalent in stochastic analysis, have been used to solve the DR problem. Recent…

Machine Learning · Computer Science 2024-10-15 Andrew Alden , Carmine Ventre , Blanka Horvath

Network sampling is used around the world for surveys of vulnerable, hard-to-reach populations including people at risk for HIV, opioid misuse, and emerging epidemics. The sampling methods include tracing social links to add new people to…

Methodology · Statistics 2020-02-05 Steve Thompson

We propose to estimate the number of communities in degree-corrected stochastic block models based on a pseudo likelihood ratio statistic. To this end, we introduce a method that combines spectral clustering with binary segmentation. This…

Methodology · Statistics 2019-07-31 Shujie Ma , Liangjun Su , Yichong Zhang