English
Related papers

Related papers: Clustering South African households based on their…

200 papers

Clustering analysis is one of the most widely used statistical tools in many emerging areas such as microarray data analysis. For microarray and other high-dimensional data, the presence of many noise variables may mask underlying…

Machine Learning · Statistics 2008-03-26 Benhuai Xie , Wei Pan , Xiaotong Shen

High-dimensional health and surveillance studies often involve many collinear predictors, multiple correlated outcomes of different types, and latent heterogeneity across observational units. We propose a Bayesian latent-cluster…

Methodology · Statistics 2026-05-13 Hsin-Hsiung Huang , Suyeon Kang

We consider a spatial version of the classical Moran model with seed-banks where the constituent populations have finite sizes. Individuals live in colonies labelled by $\mathbb{Z}^d$, $d\geq 1$, playing the role of a geographic space,…

Probability · Mathematics 2023-02-07 Frank den Hollander , Shubhamoy Nandan

Determination quadrant development has an important role in order to determine the achievement of the development of a district, in terms of the sector's gross regional domestic product (GDP). The process of determining the quadrant…

Computers and Society · Computer Science 2015-05-21 Azhari SN , Tb. Ai Munandar

Many areas of the world are without basic information on the socioeconomic well-being of the residing population due to limitations in existing data collection methods. Overhead images obtained remotely, such as from satellite or aircraft,…

Computer Vision and Pattern Recognition · Computer Science 2024-03-14 Ethan Brewer , Giovani Valdrighi , Parikshit Solunke , Joao Rulff , Yurii Piadyk , Zhonghui Lv , Jorge Poco , Claudio Silva

We analyze binary data, available for a relatively large number (big data) of families (or households), which are within small areas, from a population-based survey. Inference is required for the finite population proportion of individuals…

Methodology · Statistics 2018-06-04 Balgobin Nandram , Lu Chen , Shuting Fu , Binod Manandhar

Usual parametric and semi-parametric regression methods are inappropriate and inadequate for large clustered survival studies when the appropriate functional forms of the covariates and their interactions in hazard functions are unknown,…

Methodology · Statistics 2024-11-12 Durbadal Ghosh , Debajyoti Sinha , Antonio R. Linero , George Rust

As a quantitative characterization of the complicated economy, Macroeconomic Variables (MEVs), including GDP, inflation, unemployment, income, spending, interest rate, etc., are playing a crucial role in banks' portfolio management and…

Risk Management · Quantitative Finance 2024-05-22 Garvit Arora , Shubhangi Tiwari , Ying Wu , Xuan Mei

Segregation encodes information about society, such as social cohesion, mixing, and inequality. However, most past and current studies tackled socioeconomic (SE) segregation by analyzing static aggregated mobility networks, often without…

Current status data abounds in the field of epidemiology and public health, where the only observable data for a subject is the random inspection time and the event status at inspection. Motivated by such a current status data from a…

Methodology · Statistics 2020-04-24 Tong Wang , Kejun He , Wei Ma , Dipankar Bandyopadhyay , Samiran Sinha

Mixtures of factor analysers (MFA) models represent a popular tool for finding structure in data, particularly high-dimensional data. While in most applications the number of clusters, and especially the number of latent factors within…

Methodology · Statistics 2023-07-17 Margarita Grushanina , Sylvia Frühwirth-Schnatter

The problem of multimodal clustering arises whenever the data are gathered with several physically different sensors. Observations from different modalities are not necessarily aligned in the sense there there is no obvious way to associate…

Machine Learning · Statistics 2020-12-10 Vasil Khalidov , Florence Forbes , Radu Horaud

Clinical and epidemiological studies encode participant information in multivariate vectors with mixed type variables on continuous, truncated, ordinal, and binary scales. Semiparametric Gaussian Copula (SGC) assumes that observed data is…

Methodology · Statistics 2026-03-19 Debangan Dey , Vadim Zipunnikov

We present a novel framework for concomitant dimension reduction and clustering. This framework is based on a novel class of Bayesian clustering factor models. These models assume a factor model structure where the vectors of common factors…

Methodology · Statistics 2025-05-09 Hwasoo Shin , Marco A. R. Ferreira , Allison N. Tegge

How a household varies their regular usage of electricity is useful information for organisations to allow accurate targeting of behaviour modification initiatives with the aim of improving the overall efficiency of the electricity network.…

Computational Engineering, Finance, and Science · Computer Science 2014-09-03 Ian Dent , Tony Craig , Uwe Aickelin , Tom Rodden

We develop a structural framework for modeling and inferring unobserved heterogeneity in dynamic panel-data models. Unlike methods treating clustering as a descriptive device, we model heterogeneity as arising from a latent clustering…

Econometrics · Economics 2025-10-29 Jean-Pierre Florens , Anna Simoni

We analyse binary multivariate longitudinal data of a population of households from a rural district in South Africa. Using a 2-dimensional graphical representation of longitudinal data, each household's data is transformed into a…

Applications · Statistics 2015-11-06 Maria Vivien Visaya , David Sherwell , Charles Kimpolo , Mark Collinson

Clustering with variable selection is a challenging yet critical task for modern small-n-large-p data. Existing methods based on sparse Gaussian mixture models or sparse K-means provide solutions to continuous data. With the prevalence of…

Machine Learning · Statistics 2020-04-28 Tanbin Rahman , Yujia Li , Tianzhou Ma , Lu Tang , George Tseng

Domestic violence is commonly viewed as a gendered issue that primarily affects women, which tends to leave male victims largely overlooked. This study presents a novel, data-driven analysis of male domestic violence (MDV) in Bangladesh,…

Computers and Society · Computer Science 2025-09-26 Md Abrar Jahin , Saleh Akram Naife , Fatema Tuj Johora Lima , M. F. Mridha , Md. Jakir Hossen

The two most extended density-based approaches to clustering are surely mixture model clustering and modal clustering. In the mixture model approach, the density is represented as a mixture and clusters are associated to the different…

Machine Learning · Statistics 2016-09-16 José E. Chacón
‹ Prev 1 8 9 10 Next ›