English
Related papers

Related papers: Ancestral inference from haplotypes and mutations

200 papers

We propose a probabilistic model for interpreting gene expression levels that are observed through single-cell RNA sequencing. In the model, each cell has a low-dimensional latent representation. Additional latent variables account for…

Machine Learning · Computer Science 2018-01-18 Romain Lopez , Jeffrey Regier , Michael Cole , Michael Jordan , Nir Yosef

We consider the Moran model in continuous time with two types, mutation, and selection. We concentrate on the ancestral line and its stationary type distribution. Building on work by Fearnhead (J. Appl. Prob. 39 (2002), 38-54) and Taylor…

Populations and Evolution · Quantitative Biology 2013-12-09 Sandra Kluth , Thiemo Hustedt , Ellen Baake

We represent a process of learning by using bit strings, where 1-bits represent the knowledge acquired by individuals. Two ways of learning are considered: individual learning by trial-and-error; and social learning by copying knowledge…

Populations and Evolution · Quantitative Biology 2007-05-23 Armando Ticona Bustillos , Paulo Murilo C. de Oliveira

We develop statistically based methods to detect single nucleotide DNA mutations in next generation sequencing data. Sequencing generates counts of the number of times each base was observed at hundreds of thousands to billions of genome…

Applications · Statistics 2012-10-01 Omkar Muralidharan , Georges Natsoulis , John Bell , Hanlee Ji , Nancy R. Zhang

Several new methods have been proposed for performing valid inference after model selection. An older method is sampling splitting: use part of the data for model selection and part for inference. In this paper we revisit sample splitting…

Statistics Theory · Mathematics 2018-04-04 Alessandro Rinaldo , Larry Wasserman , Max G'Sell , Jing Lei

The prediction of phenotypic traits using high-density genomic data has many applications such as the selection of plants and animals of commercial interest; and it is expected to play an increasing role in medical diagnostics. Statistical…

Methodology · Statistics 2016-09-29 Marco Scutari , Ian Mackay , David Balding

The kernel exponential family is a rich class of distributions, which can be fit efficiently and with statistical guarantees by score matching. Being required to choose a priori a simple kernel such as the Gaussian, however, limits its…

Machine Learning · Statistics 2021-01-15 Li Wenliang , Danica J. Sutherland , Heiko Strathmann , Arthur Gretton

Most cellular phenotypes are genetically complex. Identifying the set of genes that are most closely associated with a specific cellular state is still an open question in many cases. Here we study the transcriptional profile of cellular…

Quantitative Methods · Quantitative Biology 2024-06-21 Alda Sabalic , Victoria Moiseeva , Andres Cisneros , Oleg Deryagin , Eusebio Perdiguero , Pura Muñoz-Canoves , Jordi Garcia-Ojalvo

Recently, much attention has been given to understanding recombination events along a chromosome in a variety of field. For instance, many population genetics problems are limited by the inaccuracy of inferred evolutionary histories of…

Quantitative Methods · Quantitative Biology 2017-10-31 Jacqueline Kane , Joseph Rusinko , Katherine Thompson

Likelihood-free inference involves inferring parameter values given observed data and a simulator model. The simulator is computer code which takes parameters, performs stochastic calculations, and outputs simulated data. In this work, we…

Computation · Statistics 2023-01-30 Dennis Prangle , Cecilia Viscardi

We propose a new methodology for selecting and ranking covariates associated with a variable of interest in a context of high-dimensional data under dependence but few observations. The methodology successively intertwines the clustering of…

Gaussian time-series models are often specified through their spectral density. Such models present several computational challenges, in particular because of the non-sparse nature of the covariance matrix. We derive a fast approximation of…

Computation · Statistics 2012-11-20 Nicolas Chopin , Judith Rousseau , Brunero Liseo

The observed sequence variation at a locus informs about the evolutionary history of the sample and past population size dynamics. The Kingman coalescent is used in a generative model of molecular sequence variation to infer evolutionary…

Methodology · Statistics 2021-06-22 Lorenzo Cappello , Amandine Veber , Julia A. Palacios

This paper is devoted to establishing exponential bounds for the probabilities of deviation of a sample sum from its expectation, when the variables involved in the summation are obtained by sampling in a finite population according to a…

Statistics Theory · Mathematics 2016-10-13 Patrice Bertail , Stephan Clémençon

Phylogenetic inference-the derivation of a hypothesis for the common evolutionary history of a group of species- is an active area of research at the intersection of biology, computer science, mathematics, and statistics. One assumes the…

Populations and Evolution · Quantitative Biology 2016-06-21 Ruth Davidson , Joseph Rusinko , Zoe Vernon , Jing Xi

The detection of molecular signatures of selection is one of the major concerns of modern population genetics. A widely used strategy in this context is to compare samples from several populations, and to look for genomic regions with…

Populations and Evolution · Quantitative Biology 2013-01-24 Marìa Inès Fariello , Simon Boitard , Hugo Naya , Magali SanCristobal , Bertrand Servin

In this paper we describe a new technique for the comparison of populations of DNA strands. Comparison is vital to the study of ecological systems, at both the micro and macro scales. Existing methods make use of DNA sequencing and cloning,…

Biomolecules · Quantitative Biology 2008-07-02 Dennis Shasha , Martyn Amos

We propose an adaptive importance sampling scheme for Gaussian approximations of intractable posteriors. Optimization-based approximations like variational inference can be too inaccurate while existing Monte Carlo methods can be too slow.…

Computation · Statistics 2025-02-04 Willem van den Boom , Andrea Cremaschi , Alexandre H. Thiery

Here we present the first genome wide statistical test for recessive selection. This test uses explicitly non-equilibrium demographic differences between populations to infer the mode of selection. By analyzing the transient response to a…

Populations and Evolution · Quantitative Biology 2014-03-24 Daniel J. Balick , Ron Do , David Reich , Shamil R. Sunyaev

Theoretical results for importance sampling rely on the existence of certain moments of the importance weights, which are the ratios between the proposal and target densities. In particular, a finite variance ensures square root convergence…

Methodology · Statistics 2013-07-31 Michael K. Pitt , Minh-Ngoc Tran , Marcel Scharth , Robert Kohn