English
Related papers

Related papers: Ancestral inference from haplotypes and mutations

200 papers

We provide a method for approximating Bayesian inference using rejection sampling. We not only make the process efficient, but also dramatically reduce the memory required relative to conventional methods by combining rejection sampling…

Machine Learning · Computer Science 2015-12-04 Nathan Wiebe , Christopher Granade , Ashish Kapoor , Krysta M Svore

In the case of informative sampling the sampling scheme explicitly or implicitly depends on the response variable. As a result, the sample distribution of response variable can- not be used for making inference about the population. In this…

Applications · Statistics 2016-11-18 Anna Sikov

Diagnosing an inherited disease often requires identifying the pattern of inheritance in a patient's family. We represent family trees with genetic patterns of inheritance using hypergraphs and latent state space models to provide…

Machine Learning · Statistics 2018-12-06 Edmond Cunningham , Dana Schlegel , Andrew DeOrio

Inferring dependencies between complex biological traits while accounting for evolutionary relationships between specimens is of great scientific interest yet remains infeasible when trait and specimen counts grow large. The…

Using topological summaries of gene trees as a basis for species tree inference is a promising approach to obtain acceptable speed on genomic-scale datasets, and to avoid some undesirable modeling assumptions. Here we study the…

Populations and Evolution · Quantitative Biology 2017-04-17 Elizabeth S. Allman , James H. Degnan , John A. Rhodes

Traditionally, heritability has been estimated using family-based methods such as twin studies. Advancements in molecular genomics have facilitated the development of alternative methods that utilise large samples of unrelated or related…

Recovery of population size history from molecular sequence data is an important problem in population genetics. Inference commonly relies on a coalescent model linking the population size history to genealogies. The high computational cost…

Statistics Theory · Mathematics 2018-01-17 James E. Johndrow , Julia A. Palacios

The antibody repertoire of each individual is continuously updated by the evolutionary process of B cell receptor mutation and selection. It has recently become possible to gain detailed information concerning this process through…

Populations and Evolution · Quantitative Biology 2015-05-11 Connor O. McCoy , Trevor Bedford , Vladimir N. Minin , Philip Bradley , Harlan Robins , Frederick A. Matsen

Respondent-Driven Sampling is a method to sample hard-to-reach human populations by link-tracing over their social networks. Beginning with a convenience sample, each person sampled is given a small number of uniquely identified coupons to…

Methodology · Statistics 2011-08-02 Krista J. Gile , Mark S. Handcock

Randomness is one of the important key concepts of statistics. In epidemiology or medical science, we investigate our hypotheses and interpret results through this statistical randomness. We hypothesized by imposing some conditions to this…

Methodology · Statistics 2020-02-11 T. Usuzaki , M. Shimoyama S. Chiba , S. Hotta

We apply recently developed inference methods based on general coalescent processes to DNA sequence data obtained from various marine species. Several of these species are believed to exhibit so-called shallow gene genealogies, potentially…

Populations and Evolution · Quantitative Biology 2012-11-06 Matthias Steinrücken , Matthias Birkner , Jochen Blath

To study population dynamics, ecologists and wildlife biologists use relative abundance data, which are often subject to temporal preferential sampling. Temporal preferential sampling occurs when sampling effort varies across time. To…

Methodology · Statistics 2022-12-14 Michael R. Schwob , Mevin B. Hooten , Travis McDevitt-Galles

A popular method for variance reduction in observational causal inference is propensity-based trimming, the practice of removing units with extreme propensities from the sample. This practice has theoretical grounding when the data are…

Methodology · Statistics 2024-01-30 Samir Khan , Johan Ugander

Subsequence-based time series classification algorithms provide accurate and interpretable models, but training these models is extremely computation intensive. The asymptotic time complexity of subsequence-based algorithms remains a…

Machine Learning · Computer Science 2021-02-18 Atif Raza , Stefan Kramer

Computing the exact likelihood of data in large Bayesian networks consisting of thousands of vertices is often a difficult task. When these models contain many deterministic conditional probability tables and when the observed values are…

Computation · Statistics 2012-06-26 Ydo Wexler , Dan Geiger

One of the main aims in phylogenetics is the estimation of ancestral sequences based on present-day data like, for instance, DNA alignments. One way to estimate the data of the last common ancestor of a given set of species is to first…

Populations and Evolution · Quantitative Biology 2017-02-07 Lina Herbst , Mareike Fischer

In evolutionary biology, the speciation history of living organisms is represented graphically by a phylogeny, that is, a rooted tree whose leaves correspond to current species and branchings indicate past speciation events. Phylogenies are…

Populations and Evolution · Quantitative Biology 2019-08-02 Wai-Tong Louis Fan , Sebastien Roch

We consider the problem of inference after model selection under weak assumptions in the time series setting. Even when the data are not independent, we show that sample splitting remains asymptotically valid as long as the process…

Statistics Theory · Mathematics 2019-02-27 Robert Lunde

We consider systems of slow--fast diffusions with small noise in the slow component. We construct provably logarithmic asymptotically optimal importance schemes for the estimation of rare events based on the moderate deviations principle.…

Probability · Mathematics 2020-01-07 Matthew R. Morse , Konstantinos Spiliopoulos

We introduce a methodology for performing parameter inference in high-dimensional, non-linear diffusion processes. We illustrate its applicability for obtaining insights into the evolution of and relationships between species, including…

Machine Learning · Statistics 2024-11-15 Nicklas Boserup , Gefan Yang , Michael Lind Severinsen , Christy Anna Hipsley , Stefan Sommer