中文
相关论文

相关论文: BAMITA: Bayesian Multiple Imputation for Tensor Ar…

200 篇论文

Gaussian Mixture models (GMMs) are a powerful tool for clustering, classification and density estimation when clustering structures are embedded in the data. The presence of missing values can largely impact the GMMs estimation process,…

机器学习 · 统计学 2020-06-05 Alessio Serafini , Thomas Brendan Murphy , Luca Scrucca

The problem of incomplete data - i.e., data with missing or unknown values - in multi-way arrays is ubiquitous in biomedical signal processing, network traffic analysis, bibliometrics, social network analysis, chemometrics, computer vision,…

数值分析 · 数学 2015-03-17 Evrim Acar , Tamara G. Kolda , Daniel M. Dunlavy , Morten Morup

Multiple imputation is a common approach for dealing with missing values in statistical databases. The imputer fills in missing values with draws from predictive models estimated from the observed data, resulting in multiple, completed…

统计计算 · 统计学 2018-08-30 Olanrewaju Akande , Fan Li , Jerome Reiter

To address the common problem of high dimensionality in tensor regressions, we introduce a generalized tensor random projection method that embeds high-dimensional tensor-valued covariates into low-dimensional subspaces with minimal loss of…

统计方法学 · 统计学 2025-10-03 Roberto Casarin , Radu Craiu , Qing Wang

Although approaches for handling missing data from longitudinal studies are well-developed when the patterns of missingness are monotone, fewer methods are available for non-monotone missingness. Moreover, the conventional missing at random…

统计方法学 · 统计学 2023-02-28 Boyu Ren , Stuart R. Lipsitz , Roger D. Weiss , Garrett M. Fitzmaurice

The recent development of compressed sensing has led to spectacular advances in the understanding of sparse linear estimation problems as well as in algorithms to solve them. It has also triggered a new wave of developments in the related…

信息论 · 计算机科学 2016-07-05 Christophe Schülke

Missing data in financial panels presents a critical obstacle, undermining asset-pricing models and reducing the effectiveness of investment strategies. Such panels are often inherently multi-dimensional, spanning firms, time, and financial…

应用统计 · 统计学 2025-10-09 Junyi Mo , Jiayu Li , Duo Zhang , Elynn Chen

Biomarkers are often measured in bulk to diagnose patients, monitor patient conditions, and research novel drug pathways. The measurement of these biomarkers often suffers from detection limits that result in missing and untrustworthy…

统计方法学 · 统计学 2023-06-13 Alton Barbehenn , Sihai Dave Zhao

Hierarchical data with multiple observations per group is ubiquitous in empirical sciences and is often analyzed using mixed-effects regression. In such models, Bayesian inference gives an estimate of uncertainty but is analytically…

机器学习 · 计算机科学 2026-02-05 Alex Kipnis , Marcel Binz , Eric Schulz

Imputing missing values is common practice in label-free quantitative proteomics. Imputation aims at replacing a missing value with a user-defined one. However, the imputation itself may not be optimally considered downstream of the…

统计方法学 · 统计学 2022-09-08 Marie Chion , Christine Carapito , Frédéric Bertrand

Precision medicine has led to a paradigm shift allowing the development of targeted drugs that are agnostic to the tumor location. In this context, basket trials aim to identify which tumor types - or baskets - would benefit from the…

统计方法学 · 统计学 2026-01-05 Marcio A. Diniz , Hulya Kocyigit , Erin Moshier , Madhu Mazumdar , Deukwoo Kwon

The analysis of diffusion processes in real-world propagation scenarios often involves estimating variables that are not directly observed. These hidden variables include parental relationships, the strengths of connections between nodes,…

社会与信息网络 · 计算机科学 2016-05-12 Shohreh Shaghaghian , Mark Coates

We provide a mathematical formulation and develop a computational framework for identifying multiple strains of microorganisms from mixed samples of DNA. Our method is applicable in public health domains where efficient identification of…

Factors models are routinely used to analyze high-dimensional data in both single-study and multi-study settings. Bayesian inference for such models relies on Markov Chain Monte Carlo (MCMC) methods which scale poorly as the number of…

统计方法学 · 统计学 2025-04-29 Blake Hansen , Alejandra Avalos-Pacheco , Massimiliano Russo , Roberta De Vito

Multiple imputation is widely used for handling missing data in real-world applications. For variable selection on multiply-imputed datasets, however, if selection is performed on each imputed dataset separately, it can result in different…

统计方法学 · 统计学 2025-08-07 Jungang Zou , Sijian Wang , Qixuan Chen

Mathematical models are indispensable to the system biology toolkit for studying the structure and behavior of intracellular signaling networks. A common approach to modeling is to develop a system of equations that encode the known biology…

定量方法 · 定量生物学 2024-06-18 Nathaniel Linden-Santangeli , Jin Zhang , Boris Kramer , Padmini Rangamani

Bayesian inference provides a principled framework for probabilistic reasoning. If inference is performed in two steps, uncertainty propagation plays a crucial role in accounting for all sources of uncertainty and variability. This becomes…

统计方法学 · 统计学 2026-02-16 Svenja Jedhoff , Hadi Kutabi , Anne Meyer , Paul-Christian Bürkner

Numerous studies have shown that microbial metabolites, which represent the products of bacteria in the human gut, play a key role in shaping cancer risk and response to treatment. However, metabolite data typically contain a large…

应用统计 · 统计学 2026-05-19 Kai Jiang , Satabdi Saha , Christine B. Peterson

We consider the problem of Bayesian inference for bi-variate data observed in time but with observation times which occur non-synchronously. In particular, this occurs in a wide variety of applications in finance, such as high-frequency…

统计方法学 · 统计学 2025-03-04 Ajay Jasra , Kengo Kamatani , Amin Wu

Detecting associations between microbial compositions and sample characteristics is one of the most important tasks in microbiome studies. Most of the existing methods apply univariate models to single microbial species separately, with…

统计方法学 · 统计学 2021-03-18 Boyu Ren , Sergio Bacallado , Stefano Favaro , Tommi Vatanen , Curtis Huttenhower , Lorenzo Trippa