English
Related papers

Related papers: Efficient and powerful familywise error control in…

200 papers

To understand how genetic variants in human genomes manifest in phenotypes -- traits like height or diseases like asthma -- geneticists have sequenced and measured hundreds of thousands of individuals. Geneticists use this data to build…

Machine Learning · Computer Science 2025-07-01 Alan N. Amin , Andres Potapczynski , Andrew Gordon Wilson

Thousands of risk variants underlying complex phenotypes (quantitative traits and diseases) have been identified in genome-wide association studies (GWAS). However, there are still two major challenges towards deepening our understanding of…

Methodology · Statistics 2017-10-20 Jingsi Ming , Mingwei Dai , Mingxuan Cai , Xiang Wan , Jin Liu , Can Yang

Many diseases and traits involve a complex interplay between genes and environment, generating significant interest in studying gene-environment interaction through observational data. However, for lifestyle and environmental risk factors,…

Methodology · Statistics 2023-09-22 Malka Gorfine , Conghui Qu , Ulrike Peters , Li Hsu

We develop a model-based methodology for integrating gene-set information with an experimentally-derived gene list. The methodology uses a previously reported sampling model, but takes advantage of natural constraints in the…

Methodology · Statistics 2015-06-02 Zhishi Wang , Qiuling He , Bret Larget , Michael A. Newton

We present a novel method for controlling the $k$-familywise error rate ($k$-FWER) in the linear regression setting using the knockoffs framework first introduced by Barber and Cand\`es. Our procedure, which we also refer to as knockoffs,…

Methodology · Statistics 2015-11-10 Lucas Janson , Weijie Su

In this paper, we propose the Graph-Fused Multivariate Regression (GFMR) via Total Variation regularization, a novel method for estimating the association between a one-dimensional or multidimensional array outcome and scalar predictors.…

Methodology · Statistics 2020-01-15 Ying Liu , Bowei Yan , Kathleen Merikangas , Haochang Shou

Genome-wide Association Studies (GWASs) for complex diseases often collect data on multiple correlated endo-phenotypes. Multivariate analysis of these correlated phenotypes can improve the power to detect genetic variants. Multivariate…

Methodology · Statistics 2015-03-12 Debashree Ray , James S Pankow , Saonli Basu

This paper tackles the challenge of performing multiple quantile regressions across different quantile levels and the associated problem of controlling the familywise error rate, an issue that is generally overlooked in practice. We propose…

Methodology · Statistics 2026-04-10 Riccardo De Santis , Anna Vesely , Angela Andreella

For a high-dimensional linear model with a finite number of covariates measured with error, we study statistical inference on the parameters associated with the error-prone covariates, and propose a new corrected decorrelated score test and…

Methodology · Statistics 2020-01-29 Mengyan Li , Runze Li , Yanyuan Ma

Linear mixed models (LMM) are widely adopted in genome-wide association studies (GWAS) to account for population stratification and cryptic relatedness. However, the parameter estimation of LMMs imposes substantial computational burdens due…

Computation · Statistics 2025-08-08 Zhibin Pu , Shufei Ge , Shijia Wang

Motivated by the inquiries of weak signals in underpowered genome-wide association studies (GWASs), we consider the problem of retaining true signals that are not strong enough to be individually separable from a large amount of noise. We…

Methodology · Statistics 2024-02-05 X. Jessie Jeng , Yifei Hu , Quan Sun , Yun Li

The estimation of covariance matrices of gene expressions has many applications in cancer systems biology. Many gene expression studies, however, are hampered by low sample size and it has therefore become popular to increase sample size by…

This paper develops asymptotic theory for estimation of parameters in regression models for binomial response time series where serial dependence is present through a latent process. Use of generalized linear model (GLM) estimating…

Statistics Theory · Mathematics 2016-06-06 W. T. M. Dunsmuir , J. Y. He

Motivation: The discovery of relationships between gene expression measurements and phenotypic responses is hampered by both computational and statistical impediments. Conventional statistical methods are less than ideal because they either…

Methodology · Statistics 2019-07-16 Lei Ding , Daniel J. McDonald

In genome-wide association studies (GWASs) of common diseases/traits, we often analyze multiple GWASs with the same phenotype together to discover associated genetic variants with higher power. Since it is difficult to access data with…

Genomics · Quantitative Biology 2026-03-12 Wei Jiang , Weichuan Yu

Admixture mapping is a popular tool to identify regions of the genome associated with traits in a recently admixed population. Existing methods have been developed primarily for identification of a single locus influencing a dichotomous…

Applications · Statistics 2011-11-24 Bin Zhu , Allison E. Ashley-Koch , David B. Dunson

Linear mixed models (LMMs) have emerged as the method of choice for confounded genome-wide association studies. However, the performance of LMMs in non-randomly ascertained case-control studies deteriorates with increasing sample size. We…

Genomics · Quantitative Biology 2016-02-23 Omer Weissbrod , Christoph Lippert , Dan Geiger , David Heckerman

Biological research often involves testing a growing number of null hypotheses as new data is accumulated over time. We study the problem of online control of the familywise error rate (FWER), that is testing an apriori unbounded sequence…

Methodology · Statistics 2020-03-10 Jinjin Tian , Aaditya Ramdas

In observational studies, propensity scores are commonly estimated by maxi- mum likelihood but may fail to balance high-dimensional pre-treatment covariates even after specification search. We introduce a general framework that unifies and…

Methodology · Statistics 2017-03-22 Qingyuan Zhao

Negative binomial (NB) regression is a popular method for identifying differentially expressed genes in genomics data, such as bulk and single-cell RNA sequencing data. However, NB regression makes stringent parametric and asymptotic…

Methodology · Statistics 2025-01-08 Timothy Barry , Ziang Niu , Eugene Katsevich , Xihong Lin