中文
相关论文

相关论文: Nonparametric Reduced-Rank Regression for Multi-SN…

200 篇论文

Prediction of mRNA gene-expression profiles directly from routine whole-slide images (WSIs) using deep learning models could potentially offer cost-effective and widely accessible molecular phenotyping. While such WSI-based gene-expression…

基因组学 · 定量生物学 2024-10-03 Fredrik K. Gustafsson , Mattias Rantalainen

Here we propose a test to detect effects of single nucleotide polymorphisms (SNPs) on a quantitative trait. Significant SNP-SNP interactions are more difficult to detect than significant SNPs, partly due to the massive amount of SNP-SNP…

Polygnicity refers to the phenomenon that multiple genetic variants have a non-zero effect on a complex trait. It is defined as the proportion of genetic variants that have a nonzero effect on the trait. Evaluation of polygenicity can…

基因组学 · 定量生物学 2022-07-26 Arunabha Majumdar , Bogdan Pasaniuc

Estimation of genewise variance arises from two important applications in microarray data analysis: selecting significantly differentially expressed genes and validation tests for normalization of microarray data. We approach the problem by…

统计理论 · 数学 2010-11-11 Jianqing Fan , Yang Feng , Yue S. Niu

This paper studies the problem of statistical inference for genetic relatedness between binary traits based on individual-level genome-wide association data. Specifically, under the high-dimensional logistic regression models, we define…

统计方法学 · 统计学 2022-10-06 Rong Ma , Zijian Guo , T. Tony Cai , Hongzhe Li

In the genomic era, the identification of gene signatures associated with disease is of significant interest. Such signatures are often used to predict clinical outcomes in new patients and aid clinical decision-making. However, recent…

统计方法学 · 统计学 2019-03-27 Naim U. Rashid , Quefeng Li , Jen Jen Yeh , Joseph G. Ibrahim

High resolution geospatial data are challenging because standard geostatistical models based on Gaussian processes are known to not scale to large data sizes. While progress has been made towards methods that can be computed more…

统计方法学 · 统计学 2020-12-03 Michele Peruzzi , David B. Dunson

It is now well documented that genetic covariance between functionally related traits leads to an uneven distribution of genetic variation across multivariate trait combinations, and possibly a large part of phenotype-space that is…

应用统计 · 统计学 2022-10-24 Damian Pavlyshyn , Iain M. Johnstone , Jacqueline L. Sztepanacz

We propose a scalable framework for the learning of high-dimensional parametric maps via adaptively constructed residual network (ResNet) maps between reduced bases of the inputs and outputs. When just few training data are available, it is…

A computationally simple genome-wide association study (GWAS) algorithm for estimating the main and epistatic effects of markers or single nucleotide polymorphisms (SNPs) is proposed. It is based on the intuitive assumption that changes of…

定量方法 · 定量生物学 2017-08-08 Lev V. Utkin , Irina L. Utkina

We present a novel approach for constrained Bayesian inference. Unlike current methods, our approach does not require convexity of the constraint set. We reduce the constrained variational inference to a parametric optimization over the…

机器学习 · 计算机科学 2013-09-27 Oluwasanmi Koyejo , Joydeep Ghosh

In the past several years a wide range of methods for the construction of regression trees and other estimators based on the recursive partitioning of samples have appeared in the statistics literature. Many applications involve data…

统计方法学 · 统计学 2014-07-07 Daniell Toth , John Eltinge

We propose a new structured pruning framework for compressing Deep Neural Networks (DNNs) with skip connections, based on measuring the statistical dependency of hidden layers and predicted outputs. The dependence measure defined by the…

机器学习 · 计算机科学 2022-01-28 Mohammadreza Soltani , Suya Wu , Yuerong Li , Jie Ding , Vahid Tarokh

High-dimensional k-sample comparison is a common applied problem. We construct a class of easy-to-implement nonparametric distribution-free tests based on new tools and unexplored connections with spectral graph theory. The test is shown to…

统计方法学 · 统计学 2019-08-12 Subhadeep , Mukhopadhyay , Kaijun Wang

We approach the problem of combining top-ranking association statistics or P-value from a new perspective which leads to a remarkably simple and powerful method. Statistical methods, such as the Rank Truncated Product (RTP), have been…

统计方法学 · 统计学 2019-06-12 Olga A. Vsevolozhskaya , Fengjiao Hu , Dmitri V. Zaykin

Motivation: Usefulness of analysis derived from Affymetrix microarrays depends largely upon the reliability of files describing the correspondence between probe sets, genes and transcripts. In particular, in case a gene is targeted by two…

分子网络 · 定量生物学 2012-01-16 Michel Bellis

A daunting challenge faced by modern biological sciences is finding an efficient and computationally feasible approach to deal with the curse of high dimensionality. The problem becomes even more severe when the research focus is on…

统计方法学 · 统计学 2013-04-16 Hung Hung , Yu-Tin Lin , Pengwen Chen , Chen-Chien Wang , Su-Yun Huang , Jung-Ying Tzeng

\noindent Hyper-parameter selection is a central practical problem in modern machine learning, governing regularization strength, model capacity, and robustness choices. Cross-validation is often computationally prohibitive at scale, while…

机器学习 · 统计学 2025-12-24 Hedibert Lopes , Nick Polson , Vadim Sokolov

We tackle the challenges of modeling high-dimensional data sets, particularly those with latent low-dimensional structures hidden within complex, non-linear, and noisy relationships. Our approach enables a seamless integration of concepts…

机器学习 · 统计学 2025-03-17 Zichuan Guo , Mihai Cucuringu , Alexander Y. Shestopaloff

Haplotypes, the global patterns of DNA sequence variation, have important implications for identifying complex traits. Recently, blocks of limited haplotype diversity have been discovered in human chromosomes, intensifying the research on…

基因组学 · 定量生物学 2012-07-19 Nebojsa Jojic , Vladimir Jojic , David Heckerman