English
Related papers

Related papers: Bayesian group Lasso for nonparametric varying-coe…

200 papers

Clinical investigators are increasingly interested in discovering computational biomarkers from short-term longitudinal omics data sets. This work focuses on Bayesian regression and variable selection for longitudinal omics datasets, which…

Methodology · Statistics 2025-05-20 Livia Popa , Sumanta Basu , Myung Hee Lee , Martin T. Wells

One of the major research questions regarding human microbiome studies is the feasibility of designing interventions that modulate the composition of the microbiome to promote health and cure disease. This requires extensive understanding…

Methodology · Statistics 2021-11-18 Matthew D. Koslovsky , Kristi L. Hoffman , Carrie R. Daniel , Marina Vannucci

Motivated by the inquiries of weak signals in underpowered genome-wide association studies (GWASs), we consider the problem of retaining true signals that are not strong enough to be individually separable from a large amount of noise. We…

Methodology · Statistics 2024-02-05 X. Jessie Jeng , Yifei Hu , Quan Sun , Yun Li

Genomic data arising from a genome-wide association study (GWAS) are often not only of large-scale, but also incomplete. A specific form of their incompleteness is missing values with non-ignorable missingness mechanism. The intrinsic…

Methodology · Statistics 2021-11-11 Siru Wang , Guoqi Qian

Neuroimage analysis usually involves learning thousands or even millions of variables using only a limited number of samples. In this regard, sparse models, e.g. the lasso, are applied to select the optimal features and achieve high…

Machine Learning · Computer Science 2015-03-26 Bo Xin , Lingjing Hu , Yizhou Wang , Wen Gao

Large case/control Genome-Wide Association Studies (GWAS) often include groups of related individuals with known relationships. When testing for associations at a given locus, current methods incorporate only the familial relationships…

Applications · Statistics 2014-08-01 Joshua N. Sampson , Bill Wheeler , Peng Li , Jianxin Shi

Interactions among multiple genes across the genome may contribute to the risks of many complex human diseases. Whole-genome single nucleotide polymorphisms (SNPs) data collected for many thousands of SNP markers from thousands of…

Applications · Statistics 2011-11-28 Yu Zhang , Jing Zhang , Jun S. Liu

The paper addresses joint sparsity selection in the regression coefficient matrix and the error precision (inverse covariance) matrix for high-dimensional multivariate regression models in the Bayesian paradigm. The selected sparsity…

Methodology · Statistics 2022-01-19 Srijata Samanta , Kshitij Khare , George Michailidis

Since most analysis software for genome-wide association studies (GWAS) currently exploit only unrelated individuals, there is a need for efficient applications that can handle general pedigree data or mixtures of both population and…

Applications · Statistics 2014-12-23 Hua Zhou , John Blangero , Thomas D. Dyer , Kei-hang K. Chan , Kenneth Lange , Eric M. Sobel

Gallstone disease is a complex, multifactorial condition with significant global health burdens. Identifying underlying risk factors and their interactions is crucial for early diagnosis, targeted prevention, and effective clinical…

Applications · Statistics 2025-06-18 Chitradipa Chakraborty , Nayana Mukherjee

Convolutional neural networks (CNNs) provide flexible function approximations for a wide variety of applications when the input variables are in the form of images or spatial data. Although CNNs often outperform traditional statistical…

Methodology · Statistics 2024-05-24 Yeseul Jeon , Won Chang , Seonghyun Jeong , Sanghoon Han , Jaewoo Park

In genetic association studies, detecting phenotype-genotype association is a primary goal. We assume that the relationship between the data -phenotype, genetic markers and environmental covariates - can be modelled by a generalized linear…

Methodology · Statistics 2020-04-13 K. K. Halle , Ø. Bakke , S. Djurovic , A. Bye , E. Ryeng , U. Wisløff , O. A. Andreassen , M. Langaas

In Genome-Wide Association Studies (GWAS), heritability is defined as the fraction of variance of an outcome explained by a large number of genetic predictors in a high-dimensional polygenic linear model. This work studies the asymptotic…

Statistics Theory · Mathematics 2025-02-27 David Azriel , Samuel Davenport , Armin Schwartzman

Genome-wide association studies (GWAS) have successfully identified a large number of genetic variants associated with traits and diseases. However, it still remains challenging to fully understand functional mechanisms underlying many…

Genomics · Quantitative Biology 2022-04-15 Qiaolan Deng , Jin Hyun Nam , Ayse Selen Yilmaz , Won Chang , Maciej Pietrzak , Lang Li , Hang J. Kim , Dongjun Chung

Biological data sets are often high-dimensional, noisy, and governed by complex interactions among sparse signals. This poses major challenges for interpretability and reliable feature selection. Tasks such as identifying motif interactions…

Methodology · Statistics 2025-11-20 Marta S. Lemanczyk , Lucas Kock , Johanna Schlimme , Nadja Klein , Bernhard Y. Renard

Deep feedforward neural networks (DFNNs) are a powerful tool for functional approximation. We describe flexible versions of generalized linear and generalized linear mixed models incorporating basis functions formed by a DFNN. The…

Computation · Statistics 2018-05-28 Minh-Ngoc Tran , Nghia Nguyen , David Nott , Robert Kohn

Modeling and computation for multivariate longitudinal surveys have proven challenging, particularly when data are not all continuous and Gaussian but contain discrete measurements. In many social science surveys, study participants are…

Applications · Statistics 2016-06-09 Tsuyoshi Kunihama , Carolyn T. Halpern , Amy H. Herring

The recent availability of huge, many-dimensional data sets, like those arising from genome-wide association studies (GWAS), provides many opportunities for strengthening causal inference. One popular approach is to utilize these…

Machine Learning · Statistics 2020-12-21 Ioan Gabriel Bucur , Tom Claassen , Tom Heskes

Advances of modern sensing and sequencing technologies generate a deluge of high dimensional space-temporal physiological and next-generation sequencing (NGS) data. Physiological traits are observed either as continuous random functions, or…

Machine Learning · Statistics 2014-10-28 L. Ma , N. Lin , C. I. Amos , M. M. Xiong

Important objectives in cancer research are the prediction of a patient's risk based on molecular measurements such as gene expression data and the identification of new prognostic biomarkers (e.g. genes). In clinical practice, this is…

Applications · Statistics 2020-04-17 Katrin Madjar , Manuela Zucknick , Katja Ickstadt , Jörg Rahnenführer