中文
相关论文

相关论文: Simultaneous SNP identification in association stu…

200 篇论文

In genetic association studies, detecting phenotype-genotype association is a primary goal. We assume that the relationship between the data -phenotype, genetic markers and environmental covariates - can be modelled by a generalized linear…

统计方法学 · 统计学 2020-04-13 K. K. Halle , Ø. Bakke , S. Djurovic , A. Bye , E. Ryeng , U. Wisløff , O. A. Andreassen , M. Langaas

Mediation analysis with contemporaneously observed multiple mediators is an important area of causal inference. Recent approaches for multiple mediators are often based on parametric models and thus may suffer from model misspecification.…

统计方法学 · 统计学 2022-08-30 Samrat Roy , Michael J. Daniels , Brendan J. Kelly , Jason Roy

Detecting anomalies in multivariate time series(MTS) data plays an important role in many domains. The abnormal values could indicate events, medical abnormalities,cyber-attacks, or faulty devices which if left undetected could lead to…

机器学习 · 计算机科学 2023-01-31 Usman Anjum , Samuel Lin , Justin Zhan

There is a growing interest in the estimation of the number of unseen features, mostly driven by biological applications. A recent work brought out a peculiar property of the popular completely random measures (CRMs) as prior models in…

统计方法学 · 统计学 2022-02-22 Federico Camerlenghi , Stefano Favaro , Lorenzo Masoero , Tamara Broderick

Gibbs sampling is a widely popular Markov chain Monte Carlo algorithm that can be used to analyze intractable posterior distributions associated with Bayesian hierarchical models. There are two standard versions of the Gibbs sampler: The…

统计理论 · 数学 2020-01-01 Grant Backlund , James P. Hobert , Yeun Ji Jung , Kshitij Khare

A major issue in the association of genes to neuroimaging phenotypes is the high dimension of both genetic data and neuroimaging data. In this article, we tackle the latter problem with an eye toward developing solutions that are relevant…

We expand Mendelian Randomization (MR) methodology to deal with randomly missing data on either the exposure or the outcome variable, and furthermore with data from nonindependent individuals (eg components of a family). Our method rests on…

Exploring missing data in attributed graphs introduces unique challenges beyond those found in tabular datasets. In this work, we extend the taxonomy for missing data mechanisms to attributed graphs by proposing GAMM (Graph Attributes…

机器学习 · 计算机科学 2026-02-10 Richard Serrano , Baptiste Jeudy , Charlotte Laclau , Christine Largeron

Bayesian data analysis (BDA) is today used by a multitude of research disciplines. These disciplines use BDA as a way to embrace uncertainty by using multilevel models and making use of all available information at hand. In this chapter, we…

软件工程 · 计算机科学 2020-01-03 Richard Torkar , Robert Feldt , Carlo A. Furia

As the information available to lay users through autonomous data sources continues to increase, mediators become important to ensure that the wealth of information available is tapped effectively. A key challenge that these information…

数据库 · 计算机科学 2012-08-29 Rohit Raghunathan , Sushovan De , Subbarao Kambhampati

Motivated by genetic association studies of pleiotropy, we propose here a Bayesian latent variable approach to jointly study multiple outcomes or phenotypes. The proposed method models both continuous and binary phenotypes, and it accounts…

应用统计 · 统计学 2012-11-08 Lizhen Xu , Radu V. Craiu , Lei Sun

1. Species distribution models (SDM) are tools used to determine environmental features that influence the geographic distribution of species' abundance and have been used to analyze presence-only records. Analysis of presence-only records…

种群与进化 · 定量生物学 2013-12-05 Trevor Hefley , Andrew Tyre , David Baasch , Erin Blankenship

Integrating heterogeneous datasets across different measurement platforms is a fundamental challenge in many scientific applications. A common example arises in deconvolution problems, such as cell type deconvolution, where one aims to…

统计方法学 · 统计学 2025-09-30 Dongyue Xie , Lin Gui , Jingshu Wang

Statistical dependence between hypotheses poses a significant challenge to the stability of large scale multiple hypotheses testing. Ignoring it often results in an unacceptably large spread in the false positive proportion even though the…

统计方法学 · 统计学 2018-10-15 Sairam Rayaprolu , Zhiyi Chi

As a living information and communications system, the genome encodes patterns in single nucleotide polymorphisms (SNPs) reflecting human adaption that optimizes population survival in differing environments. This paper mathematically…

种群与进化 · 定量生物学 2018-03-22 James Lindesay , Tshela E. Mason , William Hercules , Georgia M. Dunston

The original formulation of BEAMS - Bayesian Estimation Applied to Multiple Species - showed how to use a dataset contaminated by points of multiple underlying types to perform unbiased parameter estimation. An example is cosmological…

天体物理仪器与方法 · 物理学 2016-03-02 James Newling , Bruce. A. Bassett , Renée Hlozek , Martin Kunz , Mathew Smith , Melvin Varughese

Modelling gene-gene epistatic interactions when computing genetic risk scores is not a well-explored subfield of genetics and could have potential to improve risk stratification in practice. Though applications of machine learning (ML) show…

基因组学 · 定量生物学 2023-06-16 Nathaniel Gunter , Prashanthi Vemuri , Vijay Ramanan , Robel K Gebre

Combined inference for heterogeneous high-dimensional data is critical in modern biology, where clinical and various kinds of molecular data may be available from a single study. Classical genetic association studies regress a single…

应用统计 · 统计学 2017-03-22 Hélène Ruffieux , Anthony C. Davison , Jörg Hager , Irina Irincheeva

We introduce a model-based asynchronous multi-fidelity method for hyperparameter and neural architecture search that combines the strengths of asynchronous Hyperband and Gaussian process-based Bayesian optimization. At the heart of our…

机器学习 · 计算机科学 2020-07-01 Aaron Klein , Louis C. Tiao , Thibaut Lienart , Cedric Archambeau , Matthias Seeger

The Stochastic Block Model (SBM) is a popular probabilistic model for random graphs. It is commonly used for clustering network data by aggregating nodes that share similar connectivity patterns into blocks. When fitting an SBM to a network…

统计计算 · 统计学 2021-05-28 Pierre Barbillon , Julien Chiquet , Timothée Tabouy