中文
相关论文

相关论文: Clustering South African households based on their…

200 篇论文

Models for dependent data are distinguished by their targets of inference. Marginal models are useful when interest lies in quantifying associations averaged across a population of clusters. When the functional form of a covariate-outcome…

统计方法学 · 统计学 2022-04-18 Glen McGee , Alex Stringer

Modeling relations between individuals is a classical question in social sciences and clustering individuals according to the observed patterns of interactions allows to uncover a latent structure in the data. Stochastic block model (SBM)…

统计方法学 · 统计学 2015-01-27 Pierre Barbillon , Sophie Donnet , Emmanuel Lazega , Avner Bar-Hen

The Latent Dirichlet Allocation (LDA) model is a popular method for creating mixed-membership clusters. Despite having been originally developed for text analysis, LDA has been used for a wide range of other applications. We propose a new…

信息检索 · 计算机科学 2022-02-24 Gilson Shimizu , Rafael Izbicki , Denis Valle

Socioeconomic deprivation is a key determinant of public health, as highlighted by the Scottish Government's Scottish Index of Multiple Deprivation (SIMD). We propose an approach for clustering Scottish zones based on multiple deprivation…

应用统计 · 统计学 2025-08-07 Özge Şahin , Ozan Evkaya , Ariane Hanebeck

In the model-based clustering of networks, blockmodelling may be used to identify roles in the network. We identify a special case of the Stochastic Block Model (SBM) where we constrain the cluster-cluster interactions such that the density…

统计计算 · 统计学 2012-10-30 Aaron F. McDaid , Brendan Thomas Murphy , Nial Friel , Neil J. Hurley

Purpose: The primary goal of this study is to explore the application of evaluation metrics to different clustering algorithms using the data provided from the Canadian Longitudinal Study (CLSA), focusing on cognitive features. The…

机器学习 · 计算机科学 2025-05-19 ChenNingZhi Sheng

Item response theory (IRT) is a popular modeling paradigm for measuring subject latent traits and item properties according to discrete responses in tests or questionnaires. There are very limited discussions on heterogeneity pattern…

应用统计 · 统计学 2020-06-02 Guanyu Hu , Zhihua Ma , Insu Paek

We report on results concerning a partially aggregated Stock Flow Consistent (SFC) macroeconomic model in the stationary state where the sectors of banks and firms are aggregated, the sector of households is dis-aggregated, and the…

综合金融 · 定量金融 2017-08-03 Aurélien Hazan

Discrete data such as counts of microbiome taxa resulting from next-generation sequencing are routinely encountered in bioinformatics. Taxa count data in microbiome studies are typically high-dimensional, over-dispersed, and can only reveal…

统计方法学 · 统计学 2022-06-23 Yuan Fang , Sanjeena Subedi

A method for implicit variable selection in mixture of experts frameworks is proposed. We introduce a prior structure where information is taken from a set of independent covariates. Robust class membership predictors are identified using a…

计量经济学 · 经济学 2019-01-15 Gregor Zens

In the analysis of cluster data, the regression coefficients are frequently assumed to be the same across all clusters. This hampers the ability to study the varying impacts of factors on each cluster. In this paper, a semiparametric model…

统计理论 · 数学 2009-08-25 Wenyang Zhang , Jianqing Fan , Yan Sun

Massive informations about individual (household, small and medium enterprise) consumption are now provided with new metering technologies and the smart grid. Two major exploitations of these data are load profiling and forecasting at…

应用统计 · 统计学 2015-07-02 Emilie Devijver , Yannig Goude , Jean-Michel Poggi

Accurate biodiversity monitoring is essential for effective environmental policy, yet current practices often rely on arbitrarily defined ecosystems, communities, and ad-hoc indicator species, limiting cost-efficiency and reproducibility.…

应用统计 · 统计学 2025-12-02 Braden Scherting , Otso Ovaskainen , Tomas Roslin , David B. Dunson

Partially recorded data are frequently encountered in many applications and usually clustered by first removing incomplete cases or features with missing values, or by imputing missing values, followed by application of a clustering…

统计方法学 · 统计学 2021-10-20 Emily M. Goren , Ranjan Maitra

We introduce a flexible Bayesian framework for clustering nodes in undirected binary networks, motivated by the need to uncover structural patterns in complex environments. Building on the stochastic block model, we develop two hybrid…

统计方法学 · 统计学 2025-05-29 Juan Sosa , Eleni Dilma , Brenda Betancourt

We study clustering methods for binary data, first defining aggregation criteria that measure the compactness of clusters. Five new and original methods are introduced, using neighborhoods and population behavior combinatorial optimization…

Observational studies require adjustment for confounding factors that are correlated with both the treatment and outcome. In the setting where the observed variables are tabular quantities such as average income in a neighborhood, tools…

机器学习 · 统计学 2023-01-31 Connor T. Jerzak , Fredrik Johansson , Adel Daoud

With grid operators confronting rising uncertainty from renewable integration and a broader push toward electrification, Demand-Side Management (DSM) -- particularly Demand Response (DR) -- has attracted significant attention as a…

The preclinical stage of many neurodegenerative diseases can span decades before symptoms become apparent. Understanding the sequence of preclinical biomarker changes provides a critical opportunity for early diagnosis and effective…

统计方法学 · 统计学 2024-06-11 Yizhen Xu , Scott Zeger , Zheyu Wang

To understand our global progress for sustainable development and disaster risk reduction in many developing economies, two recent major initiatives - the Uniform African Exposure Dataset of the Global Earthquake Model (GEM) Foundation and…

机器学习 · 计算机科学 2026-05-28 Joshua Dimasaka , Christian Geiß , Emily So