English
Related papers

Related papers: Bayesian Hybrid Machine Learning of Gallstone Risk

200 papers

This article focuses on inference in logistic regression for high-dimensional binary outcomes. A popular approach induces dependence across the outcomes by including latent factors in the linear predictor. Bayesian approaches are useful for…

Methodology · Statistics 2025-04-23 Lorenzo Mauri , David B. Dunson

In Bayesian Networks (BNs), the direction of edges is crucial for causal reasoning and inference. However, Markov equivalence class considerations mean it is not always possible to establish edge orientations, which is why many BN structure…

Machine Learning · Computer Science 2022-10-19 Kiattikun Chobtham , Anthony C. Constantinou , Neville K. Kitson

Gaussian Graphical Models (GGMs) are widely used in high-dimensional data analysis to synthesize the interaction between variables. In many applications, such as genomics or image analysis, graphical models rely on sparsity and clustering…

Machine Learning · Statistics 2026-03-25 Do Edmond Sanou , Christophe Ambroise , Geneviève Robin

Linear mixed models (LMMs) are instrumental for regression analysis with structured dependence, such as grouped, clustered, or multilevel data. However, selection among the covariates--while accounting for this structured…

Methodology · Statistics 2022-04-20 Daniel R. Kowal

Tree-based regression and classification has become a standard tool in modern data science. Bayesian Additive Regression Trees (BART) has in particular gained wide popularity due its flexibility in dealing with interactions and non-linear…

Computation · Statistics 2022-09-13 Alan Inglis , Andrew Parnell , Catherine Hurley

The rampant adoption of ML methodologies has revealed that models are usually adopted to make decisions without taking into account the uncertainties in their predictions. More critically, they can be vulnerable to adversarial examples.…

Machine Learning · Statistics 2021-09-29 Víctor Gallego

Over the last decades, the challenges in applied regression and in predictive modeling have been changing considerably: (1) More flexible model specifications are needed as big(ger) data become available, facilitated by more powerful…

Computation · Statistics 2025-10-07 Nikolaus Umlauf , Nadja Klein , Thorsten Simon , Achim Zeileis

We develop a Bayesian non-parametric quantile panel regression model. Within each quantile, the response function is a convex combination of a linear model and a non-linear function, which we approximate using Bayesian Additive Regression…

Econometrics · Economics 2021-10-08 Todd E. Clark , Florian Huber , Gary Koop , Massimiliano Marcellino , Michael Pfarrhofer

Motivated by genome-wide association studies, we consider a standard linear model with one additional random effect in situations where many predictors have been collected on the same subjects and each predictor is analyzed separately.…

Applications · Statistics 2013-04-24 Matti Pirinen , Peter Donnelly , Chris C. A. Spencer

Motivated by the Acute Respiratory Distress Syndrome Network (ARDSNetwork) ARDS respiratory management (ARMA) trial, we developed a flexible Bayesian machine learning approach to estimate the average causal effect and heterogeneous causal…

Applications · Statistics 2024-10-29 Xinyuan Chen , Michael O. Harhay , Guangyu Tong , Fan Li

Biological data sets are often high-dimensional, noisy, and governed by complex interactions among sparse signals. This poses major challenges for interpretability and reliable feature selection. Tasks such as identifying motif interactions…

Methodology · Statistics 2025-11-20 Marta S. Lemanczyk , Lucas Kock , Johanna Schlimme , Nadja Klein , Bernhard Y. Renard

I present an application of established machine learning techniques to NHANES health survey data for predicting diabetes status. I compare baseline models (logistic regression, random forest, XGBoost) with a hybrid approach that uses an…

Machine Learning · Computer Science 2025-12-03 Mithra D K

Gene selection in high-dimensional genomic data is essential for understanding disease mechanisms and improving therapeutic outcomes. Traditional feature selection methods effectively identify predictive genes but often ignore complex…

Machine Learning · Computer Science 2025-06-02 Ehtesamul Azim , Dongjie Wang , Tae Hyun Hwang , Yanjie Fu , Wei Zhang

An important goal of environmental epidemiology is to quantify the complex health risks posed by a wide array of environmental exposures. In analyses focusing on a smaller number of exposures within a mixture, flexible models like Bayesian…

Methodology · Statistics 2024-09-27 Glen McGee , Brent A. Coull , Ander Wilson

Additive nonparametric regression models provide an attractive tool for variable selection in high dimensions when the relationship between the response and predictors is complex. They offer greater flexibility compared to parametric…

Machine Learning · Statistics 2016-07-12 Garret Vo , Debdeep Pati

Statistical methods for identifying harmful chemicals in a correlated mixture often assume linearity in exposure-response relationships. Non-monotonic relationships are increasingly recognised (e.g., for endocrine-disrupting chemicals);…

Applications · Statistics 2020-11-11 Nina Lazarevic , Luke D. Knibbs , Peter D. Sly , Adrian G. Barnett

Ecologists have long suspected that species are more likely to interact if their traits match in a particular way. For example, a pollination interaction may be more likely if the proportions of a bee's tongue fit a plant's flower shape.…

Populations and Evolution · Quantitative Biology 2019-11-05 Maximilian Pichler , Virginie Boreux , Alexandra-Maria Klein , Matthias Schleuning , Florian Hartig

Recent genome-wide association studies (GWAS) have uncovered the genetic basis of complex traits, but show an under-representation of non-European descent individuals, underscoring a critical gap in genetic research. Here, we assess whether…

Machine Learning · Computer Science 2024-05-08 Thomas Le Menestrel , Erin Craig , Robert Tibshirani , Trevor Hastie , Manuel Rivas

Understanding covariate-varying interdependencies among features is of great interest in various applications. Motivated by microbiome studies where microbial abundances and interactions vary with environmental factors, we develop a…

Methodology · Statistics 2026-03-16 Shuangjie Zhang , Michael L. Patnode , Juhee Lee

We propose a Bayesian tensor regression model to accommodate the effect of multiple factors on phenotype prediction. We adopt a set of prior distributions that resolve identifiability issues that may arise between the parameters in the…

Machine Learning · Statistics 2025-11-04 Antonia A. L. Dos Santos , Danilo A. Sarti , Rafael A. Moral , Andrew C. Parnell