English
Related papers

Related papers: The Shrinkage Variance Hotelling $T^2$ Test for Ge…

200 papers

In this paper we derive one- and two-sample multivariate empirical Bayes statistics (the $\mathit{MB}$-statistics) to rank genes in order of interest from longitudinal replicated developmental microarray time course experiments. We first…

Statistics Theory · Mathematics 2007-06-13 Yu Chuan Tai , Terence P. Speed

Benkeser et al. demonstrate how adjustment for baseline covariates in randomized trials can meaningfully improve precision for a variety of outcome types. Their findings build on a long history, starting in 1932 with R.A. Fisher and…

Methodology · Statistics 2026-03-03 Laura B. Balzer , Erica Cai , Lucas Godoy Garraza , Pracheta Amaranath

Large-scale simultaneous hypothesis testing appears in many areas such as microarray studies, genome-wide association studies, brain imaging, disease mapping and astronomical surveys. A well-known inference method is to control the false…

Methodology · Statistics 2025-07-22 Xiaoqing Niu , Pengfei Li , Yuejiao Fu

Researchers in genetics and other life sciences commonly use permutation tests to evaluate differences between groups. Permutation tests have desirable properties, including exactness if data are exchangeable, and are applicable even when…

Computation · Statistics 2018-11-01 Brian Segal , Thomas Braun , Michael Elliott , Hui Jiang

Recently-developed genotype imputation methods are a powerful tool for detecting untyped genetic variants that affect disease susceptibility in genetic association studies. However, existing imputation methods require individual-level…

Applications · Statistics 2010-11-15 Xiaoquan Wen , Matthew Stephens

Identifying disease-indicative genes is critical for deciphering disease mechanisms and has attracted significant interest in biomedical research. Spatial transcriptomics offers unprecedented insights for the detection of disease-specific…

Methodology · Statistics 2024-09-05 Qicheng Zhao , Qihuang Zhang

This paper considers inference when there is a single treated cluster and a fixed number of control clusters, a setting that is common in empirical work, especially in difference-in-differences designs. We use the t-statistic and develop…

Econometrics · Economics 2025-11-11 Chun Pong Lau , Xinran Li

We study shrinkage estimation of the mean parameters of a class of multivariate distributions for which the diagonal entries of the corresponding covariance matrix are certain quadratic functions of the mean parameter. This class of…

Statistics Theory · Mathematics 2022-07-04 Nikolas Siapoutis , Donald Richards , Bharath K. Sriperumbudur

Spatial transcriptomics (ST) is a novel technology that enables the observation of gene expression at the resolution of individual spots within pathological tissues. ST quantifies the expression of tens of thousands of genes in a tissue…

Machine Learning · Computer Science 2025-11-25 Kaito Shiku , Kazuya Nishimura , Shinnosuke Matsuo , Yasuhiro Kojima , Ryoma Bise

Permutation tests are widely used for statistical hypothesis testing when the sampling distribution of the test statistic under the null hypothesis is analytically intractable or unreliable due to finite sample sizes. One critical challenge…

Computation · Statistics 2023-08-29 Yang Shi , Huining Kang , Ji-Hyun Lee , Hui Jiang

Sharpness-Aware Minimization (SAM) was recently introduced as a regularization procedure for training deep neural networks. It simultaneously minimizes the fitness (or loss) function and the so-called fitness sharpness. The latter serves as…

Neural and Evolutionary Computing · Computer Science 2024-05-20 Illya Bakurov , Nathan Haut , Wolfgang Banzhaf

Covariate-adaptive randomization is widely employed to balance baseline covariates in interventional studies such as clinical trials and experiments in development economics. Recent years have witnessed substantial progress in inference…

Methodology · Statistics 2024-05-30 Jiahui Xin , Hanzhong Liu , Wei Ma

A key challenge in genomics is to identify genetic variants that distinguish patients with different survival time following diagnosis or treatment. While the log-rank test is widely used for this purpose, nearly all implementations of the…

Quantitative Methods · Quantitative Biology 2013-09-18 Fabio Vandin , Alexandra Papoutsaki , Benjamin J. Raphael , Eli Upfal

We study distributed algorithms implemented in a simplified biologically inspired model for stochastic spiking neural networks. We focus on tradeoffs between computation time and network complexity, along with the role of randomness in…

Neural and Evolutionary Computing · Computer Science 2017-08-22 Nancy Lynch , Cameron Musco , Merav Parter

In the genomic era, the identification of gene signatures associated with disease is of significant interest. Such signatures are often used to predict clinical outcomes in new patients and aid clinical decision-making. However, recent…

Methodology · Statistics 2019-03-27 Naim U. Rashid , Quefeng Li , Jen Jen Yeh , Joseph G. Ibrahim

In qualitative statistics, permutation tests are very popular, mainly because of their finite-sample exactness under exchangeability. However, in non-exchangeable settings, the covariance structure of permuted statistics typically differs…

Methodology · Statistics 2026-04-09 Merle Munko , Paavo Sattler

In a range of genomic applications, it is of interest to quantify the evidence that the signal at site~$i$ is active given conditionally independent replicate observations summarized by the sample mean and variance $(\bar Y, s^2)$ at each…

Statistics Theory · Mathematics 2023-12-20 Micol Tresoldi , Daniel Xiang , Peter McCullagh

While shrinkage is essential in high-dimensional settings, its use for low-dimensional regression-based prediction has been debated. It reduces variance, often leading to improved prediction accuracy. However, it also inevitably introduces…

In many transcriptomic studies, the correlation of genes might fluctuate with quantitative factors such as genetic ancestry. We propose a method that models the covariance between two variables to vary against a continuous covariate. For…

Methodology · Statistics 2021-05-03 Tae Hyun Kim , Dan Nicolae

Inference and learning for probabilistic generative networks is often very challenging and typically prevents scalability to as large networks as used for deep discriminative approaches. To obtain efficiently trainable, large-scale and well…

Machine Learning · Statistics 2017-02-08 Dennis Forster , Jörg Lücke