中文
相关论文

相关论文: Mining Hidden Populations through Attributed Searc…

200 篇论文

Respondent-driven sampling is a form of link-tracing network sampling, which is widely used to study hard-to-reach populations, often to estimate population proportions. Previous treatments of this process have used a with-replacement…

统计方法学 · 统计学 2010-06-25 Krista J. Gile

We aim at solving the problem of predicting people's ideology, or political tendency. We estimate it by using Twitter data, and formalize it as a classification problem. Ideology-detection has long been a challenging yet important problem.…

机器学习 · 计算机科学 2020-06-19 Zhiping Xiao , Weiping Song , Haoyan Xu , Zhicheng Ren , Yizhou Sun

The idea underlying the modal formulation of density-based clustering is to associate groups with the regions around the modes of the probability density function underlying the data. This correspondence between clusters and dense regions…

社会与信息网络 · 计算机科学 2021-01-22 Giovanna Menardi , Domenico De Stefano

What is a population? This review considers how a population may be defined in terms of understanding the structure of the underlying genetics of the individuals involved. The main approach is to consider statistically identifiable groups…

种群与进化 · 定量生物学 2013-06-05 Daniel John Lawson

Users of social media sites like Facebook and Twitter rely on crowdsourced content recommendation systems (e.g., Trending Topics) to retrieve important and useful information. Contents selected for recommendation indirectly give the initial…

社会与信息网络 · 计算机科学 2017-04-04 Abhijnan Chakraborty , Johnnatan Messias , Fabricio Benevenuto , Saptarshi Ghosh , Niloy Ganguly , Krishna P. Gummadi

An identity denotes the role an individual or a group plays in highly differentiated contemporary societies. In this paper, our goal is to classify Twitter users based on their role identities. We first collect a coarse-grained public…

社会与信息网络 · 计算机科学 2020-03-05 Binxuan Huang , Kathleen M. Carley

A finite set is "hidden" if its elements are not directly enumerable or if its size cannot be ascertained via a deterministic query. In public health, epidemiology, demography, ecology and intelligence analysis, researchers have developed a…

统计理论 · 数学 2019-10-17 Si Cheng , Daniel J. Eck , Forrest W. Crawford

In many networks, vertices have hidden attributes, or types, that are correlated with the networks topology. If the topology is known but these attributes are not, and if learning the attributes is costly, we need a method for choosing…

机器学习 · 统计学 2010-05-25 Xiaoran Yan , Yaojia Zhu , Jean-Baptiste Rouquier , Cristopher Moore

Information integration applications, such as mediators or mashups, that require access to information resources currently rely on users manually discovering and integrating them in the application. Manual resource discovery is a slow…

人工智能 · 计算机科学 2016-09-08 Anon Plangprasopchok , Kristina Lerman

Social networks have the surprising property of being "searchable": Ordinary people are capable of directing messages through their network of acquaintances to reach a specific but distant target person in only a few steps. We present a…

无序系统与神经网络 · 物理学 2009-11-07 D. J. Watts , P. S. Dodds , M. E. J. Newman

Feature selection can facilitate the learning of mixtures of discrete random variables as they arise, e.g. in crowdsourcing tasks. Intuitively, not all workers are equally reliable but, if the less reliable ones could be eliminated, then…

机器学习 · 统计学 2017-11-28 Vincent Zhao , Steven W. Zucker

Sentiment classification is a fundamental task in content analysis. Although deep learning has demonstrated promising performance in text classification compared with shallow models, it is still not able to train a satisfying classifier for…

人机交互 · 计算机科学 2020-04-28 Keyu Yang , Yunjun Gao , Lei Liang , Song Bian , Lu Chen , Baihua Zheng

Attribute-based person search is the task of finding person images that are best matched with a set of text attributes given as query. The main challenge of this task is the large modality gap between attributes and images. To reduce the…

计算机视觉与模式识别 · 计算机科学 2021-08-12 Boseung Jeong , Jicheol Park , Suha Kwak

Latent tree analysis seeks to model the correlations among a set of random variables using a tree of latent variables. It was proposed as an improvement to latent class analysis --- a method widely used in social sciences and medicine to…

机器学习 · 计算机科学 2016-10-04 Nevin L. Zhang , Leonard K. M. Poon

Hierarchical models are utilized in a wide variety of problems which are characterized by task hierarchies, where predictions on smaller subtasks are useful for trying to predict a final task. Typically, neural networks are first trained…

Due to a variety of reasons, such as privacy, data in the wild often misses the grouping information required for identifying minorities. On the other hand, it is known that machine learning models are only as good as the data they are…

机器学习 · 计算机科学 2025-04-22 Mohsen Dehghankar , Abolfazl Asudeh

User identification has been a major field of research in privacy and security topics. Users might utilize multiple Online Social Networks (OSNs) to access a variety of text, videos, and links, and connect to their friends. Identifying user…

社会与信息网络 · 计算机科学 2024-06-05 Yasamin Kowsari

Considering the raising socio-economic burden of autism spectrum disorder (ASD), timely and evidence-driven public policy decision making and communication of the latest guidelines pertaining to the treatment and management of the disorder…

社会与信息网络 · 计算机科学 2015-06-02 Adham Beykikhoshk , Ognjen Arandjelovic , Dinh Phung , Svetha Venkatesh , Terry Caelli

Attributed network data is becoming increasingly common across fields, as we are often equipped with information about nodes in addition to their pairwise connectivity patterns. This extra information can manifest as a classification, or as…

社会与信息网络 · 计算机科学 2018-05-22 Natalie Stanley , Marc Niethammer , Peter J. Mucha

Tabular data is prevalent across diverse domains in machine learning. With the rapid progress of deep tabular prediction methods, especially pretrained (foundation) models, there is a growing need to evaluate these methods systematically…

机器学习 · 计算机科学 2025-11-10 Han-Jia Ye , Si-Yang Liu , Hao-Run Cai , Qi-Le Zhou , De-Chuan Zhan