English
Related papers

Related papers: Improving power in genome-wide association studies…

200 papers

In applications of group testing in networks, e.g. identifying individuals who are infected by a disease spread over a network, exploiting correlation among network nodes provides fundamental opportunities in reducing the number of tests…

Information Theory · Computer Science 2023-03-21 Hesam Nikpey , Jungyeol Kim , Xingran Chen , Saswati Sarkar , Shirin Saeedi Bidokhti

We propose a general and formal statistical framework for multiple tests of association between known fixed features of a genome and unknown parameters of the distribution of variable features of this genome in a population of interest. The…

Applications · Statistics 2008-12-18 Sandrine Dudoit , Sündüz Keleş , Mark J. van der Laan

It is conventionally believed that a permutation test should ideally use all permutations. If this is computationally unaffordable, it is believed one should use the largest affordable Monte Carlo sample or (algebraic) subgroup of…

Statistics Theory · Mathematics 2023-11-27 Nick W. Koning

The group testing problem is concerned with identifying a small set of infected individuals in a large population. At our disposal is a testing procedure that allows us to test several individuals together. In an idealized setting, a test…

Information Theory · Computer Science 2023-09-19 Oliver Gebhard , Oliver Johnson , Philipp Loick , Maurice Rolvien

Motivated by the important problem of detecting association between genetic markers and binary traits in genome-wide association studies, we present a novel Bayesian model that establishes a hierarchy between markers and genes by defining…

Applications · Statistics 2016-06-22 Ian Johnston , Timothy Hancock , Hiroshi Mamitsuka , Luis Carvalho

We propose a dynamic allocation procedure that increases power and efficiency when measuring an average treatment effect in sequential randomized trials. Subjects arrive iteratively and are either randomized or paired via a matching…

Methodology · Statistics 2013-05-23 Adam Kapelner , Abba Krieger

Fitness functions based on test cases are very common in Genetic Programming (GP). This process can be assimilated to a learning task, with the inference of models from a limited number of samples. This paper is an investigation on two…

Machine Learning · Computer Science 2016-08-16 Christian Gagné , Marc Schoenauer , Marc Parizeau , Marco Tomassini

Genetic association analyses often involve data from multiple potentially-heterogeneous subgroups. The expected amount of heterogeneity can vary from modest (e.g., a typical meta-analysis) to large (e.g., a strong gene--environment…

Methodology · Statistics 2014-04-15 Xiaoquan Wen , Matthew Stephens

Gathering observational data for medical decision-making often involves uncertainties arising from both type I (false positive)and type II (false negative) errors. In this work, we develop a statistical model to study how medical…

Applications · Statistics 2025-10-21 Lucas Böttcher , Maria R. D'Orsogna , Tom Chou

Propensity score weighting is a tool for causal inference to adjust for measured confounders. Survey data are often collected under complex sampling designs such as multistage cluster sampling, which presents challenges for propensity score…

Methodology · Statistics 2016-07-27 Shu Yang

Under complete linkage disequilibrium (LD), robust tests often have greater power than Pearson's chi-square test and trend tests for the analysis of case-control genetic association studies. Robust statistics have been used in…

Methodology · Statistics 2010-10-26 Gang Zheng , Jungnam Joo , Dmitri Zaykin , Colin Wu , Nancy Geller

Learning ensembles by bagging can substantially improve the generalization performance of low-bias, high-variance estimators, including those evolved by Genetic Programming (GP). To be efficient, modern GP algorithms for evolving (bagging)…

Neural and Evolutionary Computing · Computer Science 2021-02-08 Marco Virgolin

Replicability analysis aims to identify the findings that replicated across independent studies that examine the same features. We provide powerful novel replicability analysis procedures for two studies for FWER and for FDR control on the…

Methodology · Statistics 2019-03-01 Marina Bogomolov , Ruth Heller

Randomized experiments are the "gold standard" for estimating causal effects, yet often in practice, chance imbalances exist in covariate distributions between treatment groups. If covariate data are available before units are exposed to…

Statistics Theory · Mathematics 2012-07-25 Kari Lock Morgan , Donald B. Rubin

Meta-analysis is routinely performed in many scientific disciplines. This analysis is attractive since discoveries are possible even when all the individual studies are underpowered. However, the meta-analytic discoveries may be entirely…

Methodology · Statistics 2023-05-09 Marina Bogomolov , Ruth Heller

Ensemble methods combine the predictions of multiple models to improve performance, but they require significantly higher computation costs at inference time. To avoid these costs, multiple neural networks can be combined into one by…

Machine Learning · Computer Science 2024-05-07 Alexia Jolicoeur-Martineau , Emy Gervais , Kilian Fatras , Yan Zhang , Simon Lacoste-Julien

In the coming years, advanced gravitational wave detectors will observe signals from a large number of compact binary coalescences. The majority of these signals will be relatively weak, making the precision measurement of subtle effects,…

Instrumentation and Methods for Astrophysics · Physics 2019-07-03 Aaron Zimmerman , Carl-Johan Haster , Katerina Chatziioannou

Genetic association tests involving copy-number variants (CNVs) are complicated by the fact that CNVs span multiple markers at which measurements are taken. The power of an association test at a single marker is typically low, and it is…

Methodology · Statistics 2016-07-20 Yinglei Li , Patrick Breheny

The problem of large-scale simultaneous hypothesis testing is re-visited. Bagging and subagging procedures are put forth with the purpose of improving the discovery power of the tests. The procedures are implemented in both simulated and…

Methodology · Statistics 2007-05-23 Dimitris N. Politis

Incorporating historical information into the design and analysis of a new clinical trial has been the subject of much recent discussion. For example, in the context of clinical trials of antibiotics for drug resistant infections, where…

Methodology · Statistics 2018-06-08 Isaac Gravestock , Leonhard Held