English
Related papers

Related papers: Variable selection for varying multi-index coeffic…

200 papers

Spatially varying coefficient (SVC) models are a type of regression model for spatial data where covariate effects vary over space. If there are several covariates, a natural question is which covariates have a spatially varying effect and…

Methodology · Statistics 2021-02-12 Jakob A. Dambon , Fabio Sigrist , Reinhard Furrer

The advent of high-throughput sequencing technologies has made data from DNA material readily available, leading to a surge of microbiome-related research establishing links between markers of microbiome health and specific outcomes.…

Methodology · Statistics 2018-06-19 Susheela P. Singh , Ana-Maria Staicu , Robert R. Dunn , Noah Fierer , Brian J. Reich

Many diseases and traits involve a complex interplay between genes and environment, generating significant interest in studying gene-environment interaction through observational data. However, for lifestyle and environmental risk factors,…

Methodology · Statistics 2023-09-22 Malka Gorfine , Conghui Qu , Ulrike Peters , Li Hsu

This note is concerned with an accurate and computationally efficient variational bayesian treatment of mixed-effects modelling. We focus on group studies, i.e. empirical studies that report multiple measurements acquired in multiple…

Machine Learning · Statistics 2019-03-22 Jean Daunizeau

In the genomic analysis, it is significant while challenging to identify markers associated with cancer outcomes or phenotypes. Based on the biological mechanisms of cancers and the characteristics of datasets as well, this paper proposes a…

Methodology · Statistics 2022-11-30 Yang Li , Fan Wang , Rong Li , Yifan Sun

Human migration exhibits complex spatiotemporal dependence driven by environmental and socioeconomic forces. Modeling such patterns at scale requires methods that accommodate many random effects while remaining feasible when raw data or…

Methodology · Statistics 2026-05-29 Lida Chalangar Jalili Dehkharghani , Li-Hsiang Lin

Methods utilizing instrumental variables have been a fundamental statistical approach to estimation in the presence of unmeasured confounding, usually occurring in non-randomized observational data common to fields such as economics and…

Methodology · Statistics 2022-10-06 Charles Spanbauer , Wei Pan

High-dimensional linear and nonlinear models have been extensively used to identify associations between response and explanatory variables. The variable selection problem is commonly of interest in the presence of massive and complex data.…

Methodology · Statistics 2017-08-10 Vitara Pungpapong , Min Zhang , Dabao Zhang

Gene-environment and nutrition-environment studies often involve testing of high-dimensional interactions between two sets of variables, each having potentially complex nonlinear main effects on an outcome. Construction of a valid and…

Motivated by applications in precision medicine and treatment effect heterogeneity, recent research has focused on estimating conditional average treatment effects (CATEs) using machine learning (ML). CATE estimates may represent…

Methodology · Statistics 2025-12-30 Oliver J. Hines , Karla Diaz-Ordaz , Stijn Vansteelandt

Time-to-event analyses are often plagued by both -- possibly unmeasured -- confounding and competing risks. To deal with the former, the use of instrumental variables for effect estimation is rapidly gaining ground. We show how to make use…

Methodology · Statistics 2018-01-04 Torben Martinussen , Stijn Vansteelandt

There is abundant interest in assessing the joint effects of multiple exposures on human health. This is often referred to as the mixtures problem in environmental epidemiology and toxicology. Classically, studies have examined the adverse…

Applications · Statistics 2024-05-31 Shounak Chattopadhyay , Stephanie M. Engel , David Dunson

The use of instrumental variables for estimating the effect of an exposure on an outcome is popular in econometrics, and increasingly so in epidemiology. This increasing popularity may be attributed to the natural occurrence of instrumental…

Methodology · Statistics 2016-08-03 T. Martinussen , S. Vansteelandt , E. J. Tchetgen Tchetgen , D. M. Zucker

An important goal of environmental epidemiology is to quantify the complex health risks posed by a wide array of environmental exposures. In analyses focusing on a smaller number of exposures within a mixture, flexible models like Bayesian…

Methodology · Statistics 2024-09-27 Glen McGee , Brent A. Coull , Ander Wilson

We consider nonlinear mixed effects models including high-dimensional covariates to model individual parameters variability. The objective is to identify relevant covariates among a large set under sparsity assumption and to estimate model…

Statistics Theory · Mathematics 2025-08-06 Antoine Caillebotte , Estelle Kuhn , Sarah Lemler

Statistical learning evolves quickly with more and more sophisticated models proposed to incorporate the complicated data structure from modern scientific and business problems. Varying index coefficient models extend varying coefficient…

Statistics Theory · Mathematics 2019-03-05 Li Jialiang , Lv Jing

This study introduces a nonparametric definition of interaction and provides an approach to both interaction discovery and efficient estimation of this parameter. Using stochastic shift interventions and ensemble machine learning, our…

Methodology · Statistics 2024-07-01 David B. McCoy , Alan E. Hubbard , Alejandro Schuler , Mark J. van der Laan

Mendelian randomization is an instrumental variable method that utilizes genetic information to investigate the causal effect of a modifiable exposure on an outcome. In most cases, the exposure changes over time. Understanding the…

Methodology · Statistics 2024-03-11 Haodong Tian , Ashish Patel , Stephen Burgess

It is increasingly of interest in statistical genetics to test for the presence of a mechanistic interaction between genetic (G) and environmental (E) risk factors by testing for the presence of an additive GxE interaction. In case-control…

Methodology · Statistics 2018-08-21 Eric J. Tchetgen Tchetgen , Xu Shi , Tamar Sofer , Benedict H. W. Wong

Variable selection, also known as feature selection in machine learning, plays an important role in modeling high dimensional data and is key to data-driven scientific discoveries. We consider here the problem of detecting influential…

Methodology · Statistics 2014-09-24 Bo Jiang , Jun S. Liu