English
Related papers

Related papers: Phylogenetic effective sample size

200 papers

In statistical setting of the pattern recognition problem the number of examples required to approximate an unknown labelling function is linear in the VC dimension of the target learning class. In this work we consider the question whether…

Machine Learning · Computer Science 2016-06-27 Daniil Ryabko

Phylogenetic mixture models, in which the sites in sequences undergo different substitution processes along the same or different trees, allow the description of heterogeneous evolutionary processes. As data sets consisting of longer…

Populations and Evolution · Quantitative Biology 2012-07-17 Elizabeth S. Allman , John A. Rhodes , Seth Sullivant

Phylogenetic mixture models are statistical models of character evolution allowing for heterogeneity. Each of the classes in some unknown partition of the characters may evolve by different processes, or even along different trees. The…

Populations and Evolution · Quantitative Biology 2010-11-19 John A. Rhodes , Seth Sullivant

We are interested in modelling Darwinian evolution, resulting from the interplay of phenotypic variation and natural selection through ecological interactions. Our models are rooted in the microscopic, stochastic description of a population…

Probability · Mathematics 2016-08-16 Nicolas Champagnat , Régis Ferrière , Sylvie Méléard

Feature selection (FS) is assumed to improve predictive performance and identify meaningful features in high-dimensional datasets. Surprisingly, small random subsets of features (0.02-1%) match or outperform the predictive performance of…

Machine Learning · Computer Science 2025-09-22 Bhavesh Neekhra , Debayan Gupta , Partha Pratim Chakrabarti

We consider an expanding population on the plane. The genealogy of a sample from the population is modelled by coalescing Brownian motion on the circle. We establish a weak law of large numbers for the site frequency spectrum in this model.…

Probability · Mathematics 2023-08-16 Yubo Shuai

There have been many studies to examine whether one trait is correlated with another trait across a group of present-day species (for example, do species with larger brains tend to have longer gestation times. Since the introduction of the…

Populations and Evolution · Quantitative Biology 2023-09-07 Albert Ch. Soewongsono , Barbara R. Holland , Malgorzata M. O'Reilly

The first investigation is made of designs for screening experiments where the response variable is approximated by a generalised linear model. A Bayesian information capacity criterion is defined for the selection of designs that are…

Methodology · Statistics 2016-10-27 David C. Woods , James M. McGree , Susan M. Lewis

Suppose we have a set $X$ consisting of $n$ taxa and we are given information from $k$ loci from which to construct a phylogeny for $X$. Each locus offers information for only a fraction of the taxa. The question is whether this data…

Data Structures and Algorithms · Computer Science 2020-02-25 Ghazaleh Parvini , Katherine Braught , David Fernández-Baca

Statistical consistency in phylogenetics has traditionally referred to the accuracy of estimating phylogenetic parameters for a fixed number of species as we increase the number of characters. However, as sequences are often of fixed length…

Populations and Evolution · Quantitative Biology 2010-04-09 Olivier Gascuel , Mike Steel

Phylogenetic trees are a central tool in understanding evolution. They are typically inferred from sequence data, and capture evolutionary relationships through time. It is essential to be able to compare trees from different data sources…

Populations and Evolution · Quantitative Biology 2017-10-31 Michelle Kendall , Caroline Colijn

The Central Limit Theorem provides a foundation for inferential statistics and hypothesis testing. It describes how standardized statistics behave under repeated sampling from large populations. However, if the size of the sample (n)…

Methodology · Statistics 2026-05-19 Mike Crowhurst

Phylogenetic analyses which include fossils or molecular sequences that are sampled through time require models that allow one sample to be a direct ancestor of another sample. As previously available phylogenetic inference tools assume…

Populations and Evolution · Quantitative Biology 2014-12-08 Alexandra Gavryushkina , David Welch , Tanja Stadler , Alexei Drummond

In this paper, we study the estimation of the effective number of relativistic species from a combination of CMB and BAO data. We vary different ingredients of the analysis: the Planck high-$\ell$ likelihoods, the Boltzmann solvers, and the…

Cosmology and Nongalactic Astrophysics · Physics 2019-02-27 Sophie Henrot-Versillé , Francois Couchot , Xavier Garrido , Hiroaki Imada , Thibaut Louis , Matthieu Tristram , Sylvain Vanneste

The basic idea of importance sampling is to use independent samples from a proposal measure in order to approximate expectations with respect to a target measure. It is key to understand how many samples are required in order to guarantee…

Computation · Statistics 2017-01-17 S. Agapiou , O. Papaspiliopoulos , D. Sanz-Alonso , A. M. Stuart

One essential ingredient of evolutionary theory is the concept of fitness as a measure for a species' success in its living conditions. Here, we quantify the effect of environmental fluctuations onto fitness by analytical calculations on a…

Populations and Evolution · Quantitative Biology 2015-05-19 Anna Melbinger , Massimo Vergassola

Adaptive designs have been proposed for clinical trials in which the nuisance parameters or alternative of interest are unknown or likely to be misspecified before the trial. Whereas most previous works on adaptive designs and mid-course…

Methodology · Statistics 2011-05-18 Jay Bartroff , Tze Leung Lai

Many population genetic models have been developed for the purpose of inferring population size and growth rates from random samples of genetic data. We examine two popular approaches to this problem, the coalescent and the…

Populations and Evolution · Quantitative Biology 2014-08-29 Erik M. Volz , Simon DW Frost

We present convincing empirical evidence for an effective and general strategy for building accurate small models. Such models are attractive for interpretability and also find use in resource-constrained environments. The strategy is to…

Machine Learning · Computer Science 2024-04-30 Abhishek Ghose

Accurate sample classification using transcriptomics data is crucial for advancing personalized medicine. Achieving this goal necessitates determining a suitable sample size that ensures adequate statistical power without undue resource…

Methodology · Statistics 2024-09-11 Yunhui Qi , Xinyi Wang , Li-Xuan Qin