English
Related papers

Related papers: Using prior information to boost power in correlat…

200 papers

The importance of interpretability of machine learning models has been increasing due to emerging enterprise predictive analytics, threat of data privacy, accountability of artificial intelligence in society, and so on. Piecewise linear…

Artificial Intelligence · Computer Science 2017-11-08 Masato Asahara , Ryohei Fujimaki

The need for function estimation in label-limited settings is common in the natural sciences. At the same time, prior knowledge of function values is often available in these domains. For example, data-free biophysics-based models can be…

Machine Learning · Computer Science 2022-10-17 Hunter Nisonoff , Yixin Wang , Jennifer Listgarten

Bayesian inference is attractive for its coherence and good frequentist properties. However, it is a common experience that eliciting a honest prior may be difficult and, in practice, people often take an {\em empirical Bayes} approach,…

Statistics Theory · Mathematics 2012-04-09 Sonia Petrone , Judith Rousseau , Catia Scricciolo

In pre- and non-clinical toxicology, the reduction of animal use is highly desireable. Although approaches for possible sample size reduction in the concurrent control group were suggested previously under the virtual control groups…

The problem of detecting changes in covariance for a single pair of features has been studied in some detail, but may be limited in importance or general applicability. In contrast, testing equality of covariance matrices of a {\it set} of…

Methodology · Statistics 2017-12-12 Yi-Hui Zhou

Reconstruction of a high-dimensional network may benefit substantially from the inclusion of prior knowledge on the network topology. In the case of gene interaction networks such knowledge may come for instance from pathway repositories…

Causal discovery is to learn cause-effect relationships among variables given observational data and is important for many applications. Existing causal discovery methods assume data sufficiency, which may not be the case in many real world…

Machine Learning · Computer Science 2022-06-20 Zijun Cui , Naiyu Yin , Yuru Wang , Qiang Ji

Understanding associations between paired high-dimensional longitudinal datasets is a fundamental yet challenging problem that arises across scientific domains, including longitudinal multi-omic studies. The difficulty stems from the…

Methodology · Statistics 2026-01-21 Jianbin Tan , Pixu Shi

Accurate and precise covariance matrices will be important in enabling planned cosmological surveys to detect new physics. Standard methods imply either the need for many N-body simulations in order to obtain an accurate estimate, or a…

Cosmology and Nongalactic Astrophysics · Physics 2018-12-13 Alex Hall , Andy Taylor

Research in oncology has changed the focus from histological properties of tumors in a specific organ to a specific genomic aberration potentially shared by multiple cancer types. This motivates the basket trial, which assesses the efficacy…

Applications · Statistics 2020-02-11 Jin Jin , Marie-Karelle Riviere , Xiaodong Luo , Yingwen Dong

In reliability engineering, data about failure events is often scarce. To arrive at meaningful estimates for the reliability of a system, it is therefore often necessary to also include expert information in the analysis, which is…

Methodology · Statistics 2016-10-25 Gero Walter , Frank P. A. Coolen

Covariance matrices of random vectors contain information that is crucial for modelling. Specific structures and patterns of the covariances (or correlations) may be used to justify parametric models, e.g., autoregressive models. Until now,…

Methodology · Statistics 2025-02-11 Paavo Sattler , Dennis Dobler

Important objectives in cancer research are the prediction of a patient's risk based on molecular measurements such as gene expression data and the identification of new prognostic biomarkers (e.g. genes). In clinical practice, this is…

Applications · Statistics 2020-04-17 Katrin Madjar , Manuela Zucknick , Katja Ickstadt , Jörg Rahnenführer

Strong correlation can be essentially captured with multireference wavefunction methods such as complete active space self-consistent field (CASSCF) or density matrix renormalization group (DMRG). Still, an accurate description of the…

Strongly Correlated Electrons · Physics 2022-04-06 Daria Drwal , Pavel Beran , Michał Hapka , Marcin Modrzejewski , Adam Sokół , Libor Veis , Katarzyna Pernal

Randomized A/B tests within online learning platforms represent an exciting direction in learning sciences. With minimal assumptions, they allow causal effect estimation without confounding bias and exact statistical inference even in small…

Methodology · Statistics 2023-06-13 Adam C. Sales , Ethan B. Prihar , Johann A. Gagnon-Bartsch , Neil T. Heffernan

Factor models are widely used for dimension reduction in the analysis of multivariate data. This is achieved through decomposition of a p x p covariance matrix into the sum of two components. Through a latent factor representation, they can…

Methodology · Statistics 2024-07-01 Sarah Elizabeth Heaps , Ian Hyla Jermyn

We develop a new method for frequentist multiple testing with Bayesian prior information. Our procedure finds a new set of optimal p-value weights called the Bayes weights. Prior information is relevant to many multiple testing problems.…

Methodology · Statistics 2017-10-03 Edgar Dobriban , Kristen Fortney , Stuart K. Kim , Art B. Owen

Randomized controlled trials are the gold standard for causal inference and play a pivotal role in modern evidence-based medicine. However, the sample sizes they use are often too limited to draw significant causal conclusions for subgroups…

Methodology · Statistics 2024-04-26 Xi Lin , Jens Magelund Tarp , Robin J. Evans

In this work, we offer a thorough analytical investigation into the role of shared hyperparameters in a hierarchical Bayesian model, examining their impact on information borrowing and posterior inference. Our approach is rooted in a…

Methodology · Statistics 2025-09-23 Prasenjit Ghosh , Anirban Bhattacharya , Debdeep Pati

The family-wise error rate (FWER) has been widely used in genome-wide association studies. With the increasing availability of functional genomics data, it is possible to increase the detection power by leveraging these genomic functional…

Methodology · Statistics 2020-12-25 Huijuan Zhou , Xianyang Zhang , Jun Chen