中文
相关论文

相关论文: Combining haplotypers

200 篇论文

Pattern discovery algorithms in the music domain aim to find meaningful components in musical compositions. Over the years, although many algorithms have been developed for pattern discovery in music data, it remains a challenging task. To…

声音 · 计算机科学 2020-10-26 Iris Ren , Anja Volk , Wouter Swierstra , Remco C. Veltkamp

Identifying concentrations of components from an observed mixture is a fundamental problem in signal processing. It has diverse applications in fields ranging from hyperspectral imaging to denoising biomedical sensors. This paper focuses on…

计算工程、金融与科学 · 计算机科学 2016-11-17 Shahin Mohammadi , Neta Zuckerman , Andrea Goldsmith , Ananth Grama

This paper presents a novel method to make statistical inferences for both the model support and regression coefficients in a high-dimensional logistic regression model. Our method is based on the repro samples framework, in which we…

统计方法学 · 统计学 2024-03-18 Xiaotian Hou , Linjun Zhang , Peng Wang , Min-ge Xie

Confounding matters in almost all observational studies that focus on causality. In order to eliminate bias caused by connfounders, oftentimes a substantial number of features need to be collected in the analysis. In this case, large p…

统计理论 · 数学 2019-12-30 Shinyuu Lee , Yuru Zhu

Statistical estimates can often be improved by fusion of data from several different sources. One example is so-called ensemble methods which have been successfully applied in areas such as machine learning for classification and…

物理与社会 · 物理学 2013-09-03 Johan Dahlin , Pontus Svenson

Sample overlap is a common issue in evidence synthesis in the field of medical research, particularly when integrating findings from observational studies utilizing existing databases such as registries. Due to the general inaccessibility…

统计方法学 · 统计学 2026-02-26 Zhentian Zhang , Tim Friede , Tim Mathes

This paper introduces Redescription Model Mining, a novel approach to identify interpretable patterns across two datasets that share only a subset of attributes and have no common instances. In particular, Redescription Model Mining aims to…

数据库 · 计算机科学 2021-07-12 Felix I. Stamm , Martin Becker , Markus Strohmaier , Florian Lemmerich

Persistent homology is a multiscale method for analyzing the shape of sets and functions from point cloud data arising from an unknown distribution supported on those sets. When the size of the sample is large, direct computation of the…

Computing haplotypes from sequencing data, i.e. haplotype assembly, is an important component of molecular and population genetics problems, including interpreting the effects of genetic variation on complex traits and reconstructing…

基因组学 · 定量生物学 2026-03-12 Marjan Hosseini , Ella Veiner , Thomas Bergendahl , Tala Yasenpoor , Zane Smith , Margaret Staton , Derek Aguiar

In many applications concerning statistical graphical models the data originate from several subpopulations that share similarities but have also significant differences. This raises the question of how to estimate several graphical models…

统计方法学 · 统计学 2022-06-17 Ilias Moysidis , Bing Li

Humans have $23$ pairs of homologous chromosomes. The homologous pairs are almost identical pairs of chromosomes. For the most part, differences in homologous chromosome occur at certain documented positions called single nucleotide…

信息论 · 计算机科学 2015-02-09 Govinda M. Kamath , Eren Şaşoğlu , David Tse

Machine Learning methods have of late made significant efforts to solving multidisciplinary problems in the field of cancer classification using microarray gene expression data. Feature subset selection methods can play an important role in…

计算工程、金融与科学 · 计算机科学 2013-03-04 G. Prat , Ll. Belanche

Recommender systems are established means to inspire users to watch interesting movies, discover baby names, or read books. The recommendation quality further improves by combining the results of multiple recommendation algorithms using…

信息检索 · 计算机科学 2017-10-30 Juergen Mueller

To improve the precision of inferences and reduce costs there is considerable interest in combining data from several sources such as sample surveys and administrative data. Appropriate methodology is required to ensure satisfactory…

统计方法学 · 统计学 2022-10-21 Dexter Cahoy , Joseph Sedransk

A well known problem with EOP prediction is that a prediction strategy proved to be the best for some testing period and prediction length may not remain as such for other period of time. In this paper we consider possible strategies to…

地球物理 · 物理学 2009-11-20 Zinovy Malkin

Fitting mixed models to complex survey data is a challenging problem. Most methods in the literature, including the most widely used one, require a close relationship between the model structure and the survey design. In this paper we…

统计方法学 · 统计学 2023-11-23 Thomas Lumley , Xudong Huang

A large number of approaches to Query Performance Prediction (QPP) have been proposed over the last two decades. As early as 2009, Hauff et al. [28] explored whether different QPP methods may be combined to improve prediction quality. Since…

信息检索 · 计算机科学 2025-04-01 Sourav Saha , Suchana Datta , Dwaipayan Roy , Mandar Mitra , Derek Greene

Missing data are often dealt with multiple imputation. A crucial part of the multiple imputation process is selecting sensible models to generate plausible values for incomplete data. A method based on posterior predictive checking is…

统计计算 · 统计学 2026-05-14 Mingyang Cai , Stef van Buuren , Gerko Vink

Estimation of the allele frequency at genetic markers is a key ingredient in biological and biomedical research, such as studies of human genetic variation or of the genetic etiology of heritable traits. As genetic data becomes increasingly…

应用统计 · 统计学 2007-12-18 Marc Coram , Hua Tang

Biological systems are often modelled at different levels of abstraction depending on the particular aims/resources of a study. Such different models often provide qualitatively concordant predictions over specific parametrisations, but it…

机器学习 · 统计学 2016-05-10 Giulio Caravagna , Luca Bortolussi , Guido Sanguinetti