English
Related papers

Related papers: Estimating Trees from Filtered Data: Identifiabili…

200 papers

We propose a statistical method to test whether two phylogenetic trees with given alignments are significantly incongruent. Our method compares the two distributions of phylogenetic trees given by the input alignments, instead of comparing…

Populations and Evolution · Quantitative Biology 2010-04-14 Elissaveta Arnaoudova , David Haws , Peter Huggins , Jerzy W. Jaromczyk , Neil Moore , Chris Schardl , Ruriko Yoshida

Tree shape statistics, particularly measures of tree (im)balance, play an important role in the analysis of the shape of phylogenetic trees. With applications ranging from testing evolutionary models to studying the impact of fertility…

Populations and Evolution · Quantitative Biology 2024-10-15 Sophie J. Kersting , Kristina Wicke , Mareike Fischer

Phylogenetic networks extend phylogenetic trees to allow for modeling reticulate evolutionary processes such as hybridization. They take the shape of a rooted, directed, acyclic graph, and when parameterized with evolutionary parameters,…

Populations and Evolution · Quantitative Biology 2018-08-28 R. A. L. Elworth , H. A. Ogilvie , J. Zhu , L. Nakhleh

One approach to estimating a species tree from a collection of gene trees is to first estimate probabilities of clades from the gene trees, and then to construct the species tree from the estimated clade probabilities. While a greedy…

Populations and Evolution · Quantitative Biology 2012-11-14 Elizabeth S. Allman , James H. Degnan , John A. Rhodes

Genomes and genes diversify during evolution; however, it is unclear to what extent genes still retain the relationship among species. Model species for molecular phylogenetic studies include yeasts and viruses whose genomes were sequenced…

Genomics · Quantitative Biology 2008-06-09 Yunfeng Shan , Xiu-Qing Li

Nonparametric estimation of the conditional distribution of a response given high-dimensional features is a challenging problem. It is important to allow not only the mean but also the variance and shape of the response density to change…

Machine Learning · Statistics 2013-12-05 Francesca Petralia , Joshua Vogelstein , David B. Dunson

Phylogenetic networks extend phylogenetic trees to model non-vertical inheritance, by which a lineage inherits material from multiple parents. The computational complexity of estimating phylogenetic networks from genome-wide data with…

Populations and Evolution · Quantitative Biology 2022-06-28 Jingcheng Xu , Cécile Ané

Decision trees are widely used for interpretable machine learning due to their clearly structured reasoning process. However, this structure belies a challenge we refer to as predictive equivalence: a given tree's decision boundary can be…

Machine Learning · Computer Science 2025-10-15 Hayden McTavish , Zachery Boner , Jon Donnelly , Margo Seltzer , Cynthia Rudin

The marginal likelihood of a model is a key quantity for assessing the evidence provided by the data in support of a model. The marginal likelihood is the normalizing constant for the posterior density, obtained by integrating the product…

Populations and Evolution · Quantitative Biology 2018-11-30 Mathieu Fourment , Andrew F. Magee , Chris Whidden , Arman Bilge , Frederick A. Matsen , Vladimir N. Minin

Inference of species networks from genomic data under the Network Multispecies Coalescent Model is currently severely limited by heavy computational demands. It also remains unclear how complicated networks can be for consistent inference…

Populations and Evolution · Quantitative Biology 2022-05-10 Elizabeth S. Allman , Hector Baños , Jonathan D. Mitchell , John A. Rhodes

The dynamical phenomena of complex networks are very difficult to predict from local information due to the rich microstructures and corresponding complex dynamics. On the other hands, it is a horrible job to compute some stochastic…

Data Structures and Algorithms · Computer Science 2016-01-08 Bing Yao , Xia Liu , Jin Xu

Phylogenetics uses alignments of molecular sequence data to learn about evolutionary trees. Substitutions in sequences are modelled through a continuous-time Markov process, characterised by an instantaneous rate matrix, which standard…

Populations and Evolution · Quantitative Biology 2020-07-20 Naomi E. Hannaford , Sarah E. Heaps , Tom M. W. Nye , Tom A. Williams , T. Martin Embley

We consider the task of learning Ising models when the signs of different random variables are flipped independently with possibly unequal, unknown probabilities. In this paper, we focus on the problem of robust estimation of…

Machine Learning · Statistics 2020-06-11 Ashish Katiyar , Vatsal Shah , Constantine Caramanis

Identifiability is a necessary condition for successful parameter estimation of dynamic system models. A major component of identifiability analysis is determining the identifiable parameter combinations, the functional forms for the…

Quantitative Methods · Quantitative Biology 2013-10-07 Marisa C. Eisenberg , Michael A. L. Hayashi

When hybridization or other forms of lateral gene transfer have occurred, evolutionary relationships of species are better represented by phylogenetic networks than by trees. While inference of such networks remains challenging, several…

Populations and Evolution · Quantitative Biology 2024-01-15 Elizabeth S. Allman , Hector Baños , Marina Garrote-Lopez , John A. Rhodes

Feature selection is one of the most fundamental problems in machine learning. An extensive body of work on information-theoretic feature selection exists which is based on maximizing mutual information between subsets of features and class…

Machine Learning · Statistics 2016-06-10 Shuyang Gao , Greg Ver Steeg , Aram Galstyan

Tree-based priors for probability distributions are usually specified using a predetermined, data-independent collection of candidate recursive partitions of the sample space. To characterize an unknown target density in detail over the…

Methodology · Statistics 2025-04-14 Li Ma , Benedetta Bruni

Phylogenomics heavily relies on well-curated sequence data sets that consist, for each gene, exclusively of 1:1-orthologous. Paralogs are treated as a dangerous nuisance that has to be detected and removed. We show here that this severe…

Discrete Mathematics · Computer Science 2017-12-19 Marc Hellmuth , Nicolas Wieseke , Marcus Lechner , Hans-Peter Lenhof , Martin Middendorf , Peter F. Stadler

Selective inference is considered for testing trees and edges in phylogenetic tree selection from molecular sequences. This improves the previously proposed approximately unbiased test by adjusting the selection bias when testing many trees…

Applications · Statistics 2019-05-27 Hidetoshi Shimodaira , Yoshikazu Terada

Predicting the ancestral sequences of a group of homologous sequences related by a phylogenetic tree has been the subject of many studies, and numerous methods have been proposed to this purpose. Theoretical results are available that show…

Populations and Evolution · Quantitative Biology 2013-09-05 Olivier Gascuel , Mike Steel