中文
相关论文

相关论文: Parsimonious Bayesian Factor Analysis for modellin…

200 篇论文

Bias in medical AI is often framed as a problem of representation. However, in image-based tasks such as fetal ultrasound, performance disparities can arise even when representation is adequate, because predictive accuracy depends strongly…

Multivariate spatially-oriented data sets are prevalent in the environmental and physical sciences. Scientists seek to jointly model multiple variables, each indexed by a spatial location, to capture any underlying spatial association for…

统计方法学 · 统计学 2021-08-19 Lu Zhang , Sudipto Banerjee

Factor analysis aims to determine latent factors, or traits, which summarize a given data set. Inter-battery factor analysis extends this notion to multiple views of the data. In this paper we show how a nonlinear, nonparametric version of…

机器学习 · 统计学 2016-04-19 Andreas Damianou , Neil D. Lawrence , Carl Henrik Ek

The use of a finite mixture of normal distributions in model-based clustering allows to capture non-Gaussian data clusters. However, identifying the clusters from the normal components is challenging and in general either achieved by…

统计方法学 · 统计学 2016-06-21 Gertraud Malsiner-Walli , Sylvia Frühwirth-Schnatter , Bettina Grün

The present paper proposes a novel method of quantification of the variation in biofilm architecture, in correlation with the alteration of growth conditions that include, variations of substrate and conditioning layer. The polymeric…

生物物理 · 物理学 2018-01-30 Suparna Dutta Sinha , Saptarshi Das , Sujata Tarafdar , Tapati Dutta

We introduce a methodology to construct parsimonious probabilistic models. This method makes use of Information Filtering Networks to produce a robust estimate of the global sparse inverse covariance from a simple sum of local inverse…

信息论 · 计算机科学 2017-02-21 Wolfram Barfuss , Guido Previde Massara , T. Di Matteo , Tomaso Aste

Statistical modelling strategy is the key for success in data analysis. The trade-off between flexibility and parsimony plays a vital role in statistical modelling. In clustered data analysis, in order to account for the heterogeneity…

统计方法学 · 统计学 2023-02-17 Tao Huang , Youquan Pei , Jinhong You , Wenyang Zhang

Clustering is commonly performed as an initial analysis step for uncovering structure in 'omics datasets, e.g. to discover molecular subtypes of disease. The high-throughput, high-dimensional nature of these datasets means that they provide…

统计方法学 · 统计学 2023-03-02 Paul D. W. Kirk , Filippo Pagani , Sylvia Richardson

Bi-clustering is a useful approach in analyzing biological data when observations come from heterogeneous groups and have a large number of features. We outline a general Bayesian approach in tackling bi-clustering problems in moderate to…

应用统计 · 统计学 2021-02-11 Han Yan , Jiexing Wu , Yang Li , Jun S. Liu

The analysis of clinical questionnaire data comes with many inherent challenges. These challenges include the handling of data with missing fields, as well as the overall interpretation of a dataset with many fields of different scales and…

机器学习 · 计算机科学 2021-08-04 Connor J. McLaughlin , Efi G. Kokkotou , Jean A. King , Lisa A. Conboy , Ali Yousefi

Sparse Bayesian factor models are routinely implemented for parsimonious dependence modeling and dimensionality reduction in high-dimensional applications. We provide theoretical understanding of such Bayesian procedures in terms of…

统计理论 · 数学 2014-06-03 Debdeep Pati , Anirban Bhattacharya , Natesh S. Pillai , David Dunson

Classically, Bayesian clustering interprets each component of a mixture model as a cluster. The inferred clustering posterior is highly sensitive to any inaccuracies in the kernel within each component. As this kernel is made more flexible,…

统计方法学 · 统计学 2025-12-12 David Buch , Miheer Dewaskar , David B. Dunson

We propose a Bayesian nonparametric model for mixed-type bounded data, where some variables are compositional and others are interval-bounded. Compositional variables are non-negative and sum to a given constant, such as the proportion of…

统计方法学 · 统计学 2025-03-13 Rufeng Liu , Claudia Wehrhahn , Andrés F. Barrientos , Alejandro Jara

In the analysis of cluster data, the regression coefficients are frequently assumed to be the same across all clusters. This hampers the ability to study the varying impacts of factors on each cluster. In this paper, a semiparametric model…

统计理论 · 数学 2009-08-25 Wenyang Zhang , Jianqing Fan , Yan Sun

Large-scale longitudinal molecular profiling is now firmly established in biomedical research, prompted by the need to uncover coordinated biomarker trajectories reflecting the dynamics of underlying biological mechanisms and characterise…

统计方法学 · 统计学 2026-03-24 Salima Jaoua , Daniel Temko , Hélène Ruffieux

We consider a framework for determining and estimating the conditional pairwise relationships of variables when the observed samples are contaminated with measurement error in high dimensional settings. Assuming the true underlying…

统计方法学 · 统计学 2019-07-05 Michael Byrd , Linh Nghiem , Monnie McGee

To better understand effects of exposure to food allergens, food challenge studies are designed to slowly increase the dose of an allergen delivered to allergic individuals until an objective reaction occurs. These dose-to-failure studies…

应用统计 · 统计学 2019-08-30 Matthew W. Wheeler , Joost Westerhout , Joe L. Baumert , Benjamin C. Remington

Finite mixture models have become a popular tool for clustering. Amongst other uses, they have been applied for clustering longitudinal data and clustering high-dimensional data. In the latter case, a latent Gaussian mixture model is…

统计方法学 · 统计学 2018-04-17 Vanessa S. E. Bierling , Paul D. McNicholas

Hierarchical parametric models consisting of observable and latent variables are widely used for unsupervised learning tasks. For example, a mixture model is a representative hierarchical model for clustering. From the statistical point of…

机器学习 · 统计学 2014-01-24 Keisuke Yamazaki

Data dispersed across multiple files are commonly integrated through probabilistic linkage methods, where even minimal error rates in record matching can significantly contaminate subsequent statistical analyses. In regression problems, we…

统计理论 · 数学 2024-09-18 Abhisek Chakraborty , Saptati Datta